Section

Artificial Intelligence

Models, methods, and the machinery of intelligence.

Daily at 01:15 and 13:15 UTC

Edition No. 7
Updated UTC

Top stories

  1. 1. Sharpening Tax in Post-Training

    An emerging hypothesis about reinforcement learning (RL) post-training of large language models (LLMs) is that it merely sharpens existing behaviors of a base model, improving single-shot accuracy at the cost of…

    ██████░░░░5.7
  2. 2. EvoDuet: Bilevel Co-Evolution of Web Searching and Task Solving for Scientific Discovery

    Evolutionary search with large language models (LLMs) can stall when progress requires external knowledge the model lacks.

    █████░░░░░5.5
  3. 3. Hierarchical Continuous Diffusion Language Models

    Discrete diffusion language models offer a compelling alternative to autoregressive generation for tasks demanding bidirectional reasoning and global constraint satisfaction.

    █████░░░░░5.5

More top stories

No.04 to No.08

  1. arXivby Zhang, An, Frost +5

    World Action Modeling with Progressive Visual Planning

    World action models (WAMs) have emerged as a promising paradigm for robotic control by jointly predicting future visual dynamics and actions from an initial observation and instruction.

    No.04Verdict: NotableScore█████░░░░░5.5
  2. arXivby Li, Shi, Li +7

    Scaling Trajectories for Complex Tasks through Recursive Self-Rewrite

    Successful trajectories on difficult tasks provide valuable supervision for model improvement, but specialized harnesses introduce interventions that may be unavailable during deployment.

    No.05Verdict: NotableScore█████░░░░░5.4
  3. arXivby Jiang, Yang, Ji +5

    WorldAuditBench: Interactive 3D World Auditing with Multimodal Agents

    As interactive 3D worlds are increasingly used to study intelligent behavior, it becomes important to develop efficient pipelines for identifying anomalies in these simulated ....

    No.07Verdict: RoutineScore█████░░░░░5.4

Briefs

10 more from the desk