返回首页
a
arXiv cs.AI热榜
10月3日 11:20 更新,当前榜单前三是「Heavy-Tailed Memory…」。「When Do Causal Worl…」。「From Proposal to Ve…」。这是一个人工智能方向的信息源,内容以模型进展、研究与工具发布为主。本页当前收录 50 条,保留arXiv cs.AI的原始排序和指标,不把不同平台的数据直接相加,每条标题尽量直达原始页面。热摸爽只做聚合与排序,内容版权归arXiv cs.AI及原作者所有。
也常被搜索为:arXiv cs.AI热点
1
Heavy-Tailed Memory Traces in Long-Horizon Language Agents
2 3790.5 万热度发布于 23小时前 抓取于 刚刚
When Do Causal World Models Help Modular LLM Agents
3 3622.4 万热度发布于 23小时前 抓取于 刚刚
From Proposal to Verified Effect: Praxa, an Evidence-Bound Harness for Governed AI Agent Execution
4 3458.5 万热度发布于 23小时前 抓取于 刚刚
What Do Rationales Communicate? A Message-Intervention Study in Role-Specialized QA
5 3298.7 万热度发布于 23小时前 抓取于 刚刚
Measuring the Microtask Eligibility Gap: When Is an Off-the-Shelf SLM Enough for an Agent Harness?
6 3142.9 万热度发布于 23小时前 抓取于 刚刚
Characterizing a Configuration Where Inference-Time PRM-Pruned Fragment Grafting Is Inert: Evidence from Three Reasoning LMs
7 2991.3 万热度发布于 23小时前 抓取于 刚刚
Gradient-Aligned Pair Selection for Personalized Preference Optimization
8 2843.8 万热度发布于 23小时前 抓取于 刚刚
K-Dense BYOK: An Open-Source AI Research Assistant That Runs Locally and Keeps a Hash-Chained Lab Notebook
9 2700.3 万热度发布于 23小时前 抓取于 刚刚
Scientific Agents: Evaluating Profession-Specific System Prompts on Scientific Tasks
10 2560.9 万热度发布于 23小时前 抓取于 刚刚
Comedic Fool's Gold: Reward Exploits and Countermeasures in Conversational Humor
11 2425.5 万热度发布于 23小时前 抓取于 刚刚
EviGraph: Proof-Carrying Selective Recommendation over Temporal Public-Service Knowledge Graphs
12 2294.1 万热度发布于 23小时前 抓取于 刚刚
Build2SPARQL: A Large-Scale Text-to-SPARQL Benchmark Dataset for Building Knowledge Graph Querying
13 2166.7 万热度发布于 23小时前 抓取于 刚刚
Robust Is Salient: An Informed Adversary Moves the Optimal Signal onto the Salience Pole
14 2043.2 万热度发布于 23小时前 抓取于 刚刚
Conflicting Supervision Moves Commitment, Not Capability: A 12.29{\sigma} arrangement effect that is exactly zero under a convention-agnostic score
15 1923.7 万热度发布于 23小时前 抓取于 刚刚
Knowing When to Yield: Grounded Arbitration of User Corrections in Text-Based Embodied Agents
16 1808.1 万热度发布于 23小时前 抓取于 刚刚
Rules to Tools: Executable Checks for LLM Agents in Scientific Computing
17 1696.4 万热度发布于 23小时前 抓取于 刚刚
Predictive Credit: Measuring What Scientific Explanations Add to Experimental Forecasts
18 1588.6 万热度发布于 23小时前 抓取于 刚刚
ContractRL: Shielded Group-Relative Policy Optimization for Auditable Tool-Call Repair
19 1484.6 万热度发布于 23小时前 抓取于 刚刚
Mathematical Transfer in LLMs Follows Reasoning Approach More Than Topic
20 1384.4 万热度发布于 23小时前 抓取于 刚刚
Fault-Tolerant Budget Conservation in Distributed Multi-Agent Delegation
1288.1 万热度发布于 23小时前 抓取于 刚刚