Roller | AI AGENT 中文社区 - Telegram Channel
Publisher updates
Follow the original publisher. These items are discovery material—not automatically approved reporting, endorsements, or verified breaking news from TLB.
Collection recently observed.Source timestamps describe the item, not necessarily a new event. 5650 stored items match this view.
arXiv AI
Probabilistic Focal Search: Accelerating Bounded-Suboptimal Search via Lower-Bound Advancement
External source · read at the original publisherarXiv AI
Automating Quadratic Unconstrained Binary Optimization (QUBO) Formulation Generation from Natural Language
External source · read at the original publisherarXiv AI
A Multi-Stage Rule-Chaining Framework for Compositional and Interpretable Cognitive Reasoning
External source · read at the original publisherarXiv AI
Understanding LoRA Rank Trade-offs in Diffusion Model Fine-Tuning
External source · read at the original publisherarXiv AI
Quantifying the Memorization-to-Generalization Transition: Scaling Laws and Phase Structure in Grokking
External source · read at the original publisherarXiv AI
An Open Recipe for IMO Gold: Training Nemotron for Olympiad Mathematics
External source · read at the original publisherarXiv AI
Finishing the Task Is Not Enough: Evaluating Agent Resilience and Considerate Participation under Accumulating Challenge
External source · read at the original publisherarXiv AI
Towards a Deterministic Math Solver for Clinical Language Models
External source · read at the original publisherarXiv AI
Studying Without a Syllabus: Task-Agnostic Environment Preprocessing
External source · read at the original publisherarXiv AI
When Validation Stops Learning: Auditing Update Admission for Continual Embodied Agents
External source · read at the original publisherarXiv AI
Decoupling Readiness from Release for Tail-Aware Scheduling of Agentic LLM Workflows
External source · read at the original publisherarXiv AI
Demystifying the Privacy-Utility Trade-off in LLM Interactions
External source · read at the original publisherarXiv AI
Defining AI Agents: A Compendium of Criteria, Metrics, and Benchmarks
External source · read at the original publisherarXiv AI
The Agent Incident Registry: Toward Preventing Repeated AI Agent Failures
External source · read at the original publisherarXiv AI
Grounding Agent Memory: Environment-Probing Curation for Enterprise Agents
External source · read at the original publisherarXiv AI
Fork Where the Model Changes Its Mind: Belief-Shift Branching for Tree-Structured Reinforcement Learning
External source · read at the original publisherarXiv AI
MOSAIC: Query-Aware Exploration Policy Adaptation for GraphRAG
External source · read at the original publisherarXiv AI
Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation
External source · read at the original publisherarXiv AI
KuaiRP Series Role-playing Models Technical Report
External source · read at the original publisherarXiv AI
Same Day, Same Story; One Day Ahead, a Different Signal: The Dual Validity of Financial Sentiment
External source · read at the original publisherarXiv AI
The Oligarch Barely Steers Model Collapse in Multi-Model Ecosystems
External source · read at the original publisherarXiv AI
Autonomous Chemical Mechanistic Discovery through Agentic Reasoning and Validation
External source · read at the original publisherarXiv AI
DRG-MAPPO: Hierarchical Dynamic Role-Graph Multi-Agent Reinforcement Learning for Cooperative Air Combat
External source · read at the original publisherarXiv AI
Breaking Predictions Is Not Enough: Specified-Foil Counterfactuals for Temporal Graphs
External source · read at the original publisherarXiv AI
Debate-to-Skill: Capability-Bound Process Supervision for Industrial Query-to-Agent Annotation
External source · read at the original publisherarXiv AI
SemVerBench: Benchmarking LLM Comprehension of Version-Constraint Resolution Semantics
External source · read at the original publisherarXiv AI
Can LLMs Follow Medical Expert Logic? A Benchmark for Hierarchical Logical Consistency in Risk-of-Bias Assessment
External source · read at the original publisherarXiv AI
Agentic Share-of-Search: A Multi-Agent AI System for Competitive Decision-Making in LLM-Mediated E-Commerce
External source · read at the original publisherarXiv AI