AI Research Brief
Search
Methodology
中文
The Model as Its Own Teacher, and Why Splitting Crowds Wins
7 selected from 200 papers
Featured
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning
score 10
入选 HF Daily Papers; HF 热度: 85 upvotes (+4); 有代码实现; 关键词(3): distillation, agentic, tool use
SearchOS-V1: Towards Robust Open-Domain Information-Seeking Agent Collaboration
score 10
入选 HF Daily Papers; HF 热度: 59 upvotes (+4); 有代码实现; 关键词(1): throughput
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination
score 10
入选 HF Daily Papers; HF 热度: 23 upvotes (+4); 有代码实现; 关键词(4): pretraining, reasoning, vision-language, embodied
Partition, Prompt, Aggregate: Statistical Self-Consistency in Language Models
score 7
入选 HF Daily Papers; HF 热度: 7 upvotes (+2); 有代码实现
Chat2Scenic: An Iterative RAG-Based Framework for Scenario Generation in Autonomous Driving
score 7
入选 HF Daily Papers; HF 热度: 4 upvotes (+1); 有代码实现; 关键词(3): retrieval-augmented, RAG, open source
Smarter and Cheaper at Once: Byte-Exact KV-Cache Grafting Turns a Frozen Small Model into a Verified-Knowledge Flywheel
score 6
入选 HF Daily Papers; HF 热度: 10 upvotes (+3)
Also Worth Noting
Token Time Continuous Diffusion for Language Modeling
score 5
入选 HF Daily Papers; HF 热度: 7 upvotes (+2)