TechPulseTelegram
← PrevWed, Aug 26Next β†’

πŸ€– AI Agent7 items

Latest AI agent research papers, analysis, and improvement insights

semantic_scholarresearch

The Compaction Cliff in Long-Running AI Agent Memory

μ»¨ν…μŠ€νŠΈ μ••μΆ• μ‹œ μ•ˆμ „ κ·œμΉ™μ΄ μ†μ‹€λ˜λŠ” 문제λ₯Ό λ°œκ²¬ν•˜κ³  Knowledge Triage둜 ν•΄κ²°

semantic_scholarresearch

ExecRubrics: Executable Tool-Augmented Rubrics for Verifiable and Efficient Long-Form Evaluation

Tool 기반 LLM μ—μ΄μ „νŠΈμ˜ 결과물을 ν”„λ‘œκ·Έλž˜λ§€ν‹±ν•˜κ²Œ 검증 ν‰κ°€ν•˜λŠ” 방식

semantic_scholarresearch

Forgotten in Weights, Recovered by Tools: Agentic Tool Unlearning for LLM Agents

Unlearning 후에도 도ꡬλ₯Ό 톡해 정보 λ³΅κ΅¬λ˜λŠ” 문제점과 λŒ€μ‘ λ°©μ•ˆ 연ꡬ

semantic_scholarresearch

From Language Models to Agentic AI: A Survey of Autonomous, Action-Enabled, and Collaborative LLM Agents

μ—μ΄μ „νŠΈ μ•„ν‚€ν…μ²˜, 도ꡬ ν™œμš©, ν˜‘μ—… 방식과 평가 ν‘œμ€€μ„ μ’…ν•©ν•œ μ΅œμ‹  μ„œλ² μ΄

semantic_scholarresearch

AI Watchdog: Agent Interfaces for Detecting and Defending Against Manipulative Dark Patterns in AI Conversations

AI λŒ€ν™”μ˜ 5κ°€μ§€ μ‘°μž‘ νŒ¨ν„΄μ„ μ‹€μ‹œκ°„ νƒμ§€ν•˜λŠ” λΈŒλΌμš°μ € 기반 보호 μ‹œμŠ€ν…œ

semantic_scholarresearch

Who Delegates to AI? Evidence from 53,000 Agent Configurations

53,000개 μ—μ΄μ „νŠΈ λΆ„μ„μœΌλ‘œ 직업별 AI λ„μž… ν˜„ν™©κ³Ό 에이전틱 채택 μ§€μˆ˜ λ„μΆœ

semantic_scholarresearch

Designing a Robust LLM-Based Evaluation System for Agentic AI in Drug Discovery Through Human Alignment

μ•½λ¬Ό 발견 λ„λ©”μΈμ—μ„œ Tool 기반 μ—μ΄μ „νŠΈμ˜ μ„±λŠ₯을 인간 μ •λ ¬μœΌλ‘œ 평가