Expand arXiv paper corpus
This commit is contained in:
@@ -0,0 +1,61 @@
|
||||
# Paper: Diagnosing Task Insensitivity in Language Agents
|
||||
|
||||
---
|
||||
type: paper
|
||||
title: Diagnosing Task Insensitivity in Language Agents
|
||||
authors: Jingyu Liu, Xiaopeng Wu, Kehan Chen, Chuan Yu, Yong Liu
|
||||
year: 2026
|
||||
venue: arXiv
|
||||
url: https://arxiv.org/abs/2606.26918
|
||||
code_url:
|
||||
source: arxiv
|
||||
collected_at: 2026-07-08
|
||||
published_at: 2026-06-25
|
||||
updated_at: 2026-06-25
|
||||
status: queued
|
||||
relevance: high
|
||||
topics:
|
||||
- agent-evaluation
|
||||
- planning
|
||||
- rag
|
||||
- tool-use
|
||||
methods:
|
||||
-
|
||||
benchmarks:
|
||||
-
|
||||
models:
|
||||
-
|
||||
datasets:
|
||||
- cs.AI
|
||||
related_concepts:
|
||||
-
|
||||
related_jobs:
|
||||
-
|
||||
related_experiments:
|
||||
-
|
||||
related_projects:
|
||||
-
|
||||
collection_score: 15
|
||||
collection_queries: language-agent
|
||||
---
|
||||
|
||||
## One-line Takeaway
|
||||
|
||||
Auto-collected from arXiv because it matched the Agent collection queries. Needs human skim.
|
||||
|
||||
## Why Collected
|
||||
|
||||
- matched queries: language-agent
|
||||
- inferred topics: agent-evaluation, planning, rag, tool-use
|
||||
- arXiv categories: cs.AI
|
||||
- collection score: 15
|
||||
|
||||
## Review Checklist
|
||||
|
||||
- Does this paper directly inform Agent architecture, evaluation, memory, tools, safety, coding agents, GUI/browser agents, or multi-agent workflows?
|
||||
- Does it include a benchmark, dataset, code, or reproducible experimental setup?
|
||||
- Should it be promoted from `queued` to `skimmed` or `summarized`?
|
||||
|
||||
## Links
|
||||
|
||||
- arXiv: https://arxiv.org/abs/2606.26918
|
||||
Reference in New Issue
Block a user