返回首页
a
arXiv cs.AI热榜
10月10日 11:15 更新,当前榜单前三是「An Explainable Head…」。「Synthesis Through S…」。「Agent-Controlled Fo…」。这是一个人工智能方向的信息源,内容以模型进展、研究与工具发布为主。本页当前收录 50 条,保留arXiv cs.AI的原始排序和指标,不把不同平台的数据直接相加,每条标题尽量直达原始页面。本站只做聚合与排序,内容版权归arXiv cs.AI及原作者所有。
也常被搜索为:arXiv cs.AI热点
1
An Explainable Header-Centric Framework for Large-Scale Semantic Table Interpretation and Data Quality Assessment
2 3792.9 万热度发布于 23小时前 抓取于 刚刚
Synthesis Through Simulation: Generating Coherent Enterprise Data via Scalable Agent-System Interaction
3 3624.7 万热度发布于 23小时前 抓取于 刚刚
Agent-Controlled Forgetting for Tool-Using Agents: Reversible Context Curation in Practice
4 3460.6 万热度发布于 23小时前 抓取于 刚刚
Verification and Self-Improvement in Agentic AI: Foundations and Limits
5 3300.7 万热度发布于 23小时前 抓取于 刚刚
The Harness as the Only Mutable Surface: Compliance-Bounded Self-Evolution of LLM Agents in Credit Pipelines, with a Measured Admission Gate
6 3144.9 万热度发布于 23小时前 抓取于 刚刚
Speaking the Navigator's Language: Trajectory-Grounded Instruction Translation for Frozen Aerial VLN Agents
7 2993.2 万热度发布于 23小时前 抓取于 刚刚
Plan-and-Patch: Diffusion Language Models for Agentic Planning
8 2845.6 万热度发布于 23小时前 抓取于 刚刚
Whose Ground Truth? Embracing Ambiguity in Human-Centered AI
9 2702 万热度发布于 23小时前 抓取于 刚刚
On the Clock: Towards Punctual and Productive Time-Budgeted AI Agents
10 2562.5 万热度发布于 23小时前 抓取于 刚刚
Self-Supervised Keyframe Discovery for Horizon-Invariant Behavior Cloning
11 2427 万热度发布于 23小时前 抓取于 刚刚
Reading the Room: Foundations, Design, and Challenges of Normative Competence in LLMs
12 2295.5 万热度发布于 23小时前 抓取于 刚刚
StoreBench: A Live-Commerce Environment for Evaluating and Training Autonomous Operator Agents
13 2168 万热度发布于 23小时前 抓取于 刚刚
Learning How to Search for Plans with Exponentially Less Space
14 2044.5 万热度发布于 23小时前 抓取于 刚刚
How Narrative Wrapping Affects LLM Refusal: A Cross-Language Benchmark and Defense
15 1924.9 万热度发布于 23小时前 抓取于 刚刚
Curating Always-Loaded Context for LLM Agents: A Capacitated Assortment Model with Censored Feedback
16 1809.2 万热度发布于 23小时前 抓取于 刚刚
Distillation for Incrimination and Distillation for Capabilities
17 1697.5 万热度发布于 23小时前 抓取于 刚刚
AgentHorizon: Evaluating Agentic Judges for Long-Horizon Computer-Use Tasks
18 1589.6 万热度发布于 23小时前 抓取于 刚刚
Beyond Imitation: A Framework and Benchmark for LLM-Assisted Peer Review
19 1485.5 万热度发布于 23小时前 抓取于 刚刚
OpenProblemBench: Benchmarking AI on Open Problems in the Foundational Theoretical Sciences
20 1385.3 万热度发布于 23小时前 抓取于 刚刚
When Interfaces Speak: Data-Aware Generative UI Harness for Active Interaction
1288.9 万热度发布于 23小时前 抓取于 刚刚