IndicBankBench: Evaluating Safety and Reliability of Language Model Assistants in Indian Retail Banking Paper • 2609.29167 • Published 7 days ago • 18
PUBG Ally: A Conversational Embodied Agent as an AI Teammate Paper • 2609.29837 • Published 7 days ago • 25
Just Ask Jev: Reinforcement Learning for Calibrated Decisions as a Zero-Shot Detector of AI Alignment Failures Paper • 2609.29429 • Published 7 days ago • 23
Agent-Editing World Model: Rethinking World Modeling for LLM Agents Paper • 2609.28416 • Published 8 days ago • 41
Just-in-Time Memory: Learning to Curate Task-Adaptive Memory for LLM Agents Paper • 2609.27334 • Published 8 days ago • 51
SpeakerMem-R1: Speaker-Centered Dual-Track Memory for Multi-Party Dialogue Paper • 2609.26780 • Published 9 days ago • 101
lmstudio-community/Llama-3-Groq-8B-Tool-Use-GGUF Text Generation • 8B • Updated Jul 18, 2024 • 2k • 23
SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness Paper • 2609.20519 • Published 14 days ago • 138
ImIR: Image-Instruction Tuning for All-in-One Image Restoration Paper • 2609.25267 • Published 10 days ago • 14