← back to terminalTYPE0//PAPERS

breaking papers · 78 analyzed

The most important papers, decoded.

AI-powered analysis of breakthrough research from arXiv and beyond. We surface the work that matters before it hits the news cycle.

  • arXiv:2512.01031·5h 59m ago

    A new AI lets robots 'think ahead' and finish tasks up to twice as fast

    MIT's robotics AI system VLASH plans a robot's next move while the current one runs, cutting reaction delays up to 11.8× on the same hardware. Demos are research benchmarks, not deployed robots.

    →
  • arXiv:2608.09119·8h 44m ago

    Korea's Motif 3 Just Went From Research-Only to Commercial-Ready

    Motif Technologies re-released its largest open-weights language model under an MIT license, letting any builder fine-tune, embed, or sell products built on the weights.

    →
  • arXiv:2511.23404·8h 45m ago

    Liquid AI's Vision Model Lives on a Phone Because Its Working Memory Stays Small

    Liquid AI's 3.1B-parameter vision-language model (an AI that reads images and answers questions about them) swaps the standard AI architecture's growing memory buffer for a fixed-size state, fitting in 3.

    →
  • arXiv:2512.15889·10h 1m ago

    Xanadu and University of Alberta will tackle a specific bottleneck in light-activated cancer drug chemistry

    A public photonic-quantum company and University of Alberta chemists are pairing quantum algorithms designed for future error-correcting hardware with classical benchmarks on light-activated cancer drug molecules (photosensitizers).

    →
  • arXiv:2608.12306·10h 16m ago

    A cheaper way to teach robots and self-driving systems not to crash

    Researchers show that a single 'stop' label per unsafe moment can teach the whole trajectory, by redistributing the safety signal backward through every earlier action.

    →
  • arXiv:2411.04872·13h 20m ago

    Anthropic researcher and Claude build every open Hadamard matrix below size 2000

    Anthropic researcher Levent Alpöge, with two co-authors and Claude, closed the smallest open Hadamard matrix, a +1/−1 array stuck since 2005, and filled 11 more below 2000, with AI benchmarking group Epoch AI's mark provisional.

    →
  • arXiv:2607.11063·14h 22m ago

    A smudge on the camera can send a delivery robot to the wrong floor. Researchers just proved it.

    A black-box adversarial attack — one that never needs the robot's model internals — from the Hong Kong University of Science and Technology (Guangzhou), accepted at the major computer science conference ACM Multimedia 2026, finds the most modern

    →
  • arXiv:2608.11573·15h 8m ago

    Researchers train LLMs to check each step of their own reasoning

    A new framework called SFS-DPO (Self-Fix Step-DPO) splits step-level reasoning from self-verification, with reported gains on math and code benchmarks over prior step-level training methods.

    →
  • arXiv:2608.11658·15h 11m ago

    When AI Agents Work in Teams, Letting Each One Pick Its Own Best Move Can Backfire

    A new arXiv paper proves that cooperative AI systems, teams of agents sharing one goal, can collectively pick worse moves than any in their shared playbook.

    →
  • arXiv:2608.12246·16h 30m ago

    When Tools Hunt for the Code Change That Introduced a Bug, They Get It Wrong Two-Thirds of the Time

    A 100-code-change test across Python, Java, and C++ shows why state-of-the-art tools that try to pin down where a bug entered the code still need human auditors: the code changes that introduced the flaws run about six times larger than the fixes

    →
  • arXiv:2507.10463·16h 48m ago

    The Chip Industry Is Finally Stopping Designing Hardware and Software Separately

    AI inference and thermal ceilings have made the old hardware-to-software handoff uneconomic, spurring a second attempt 30 years on to design chips and the software that runs on them as one system, with virtual-twin simulations — software models of

    →
  • arXiv:2608.11357·18h 13m ago

    When More Capable AI Agents Make a Worse Team

    A new preprint on multi-agent AI systems argues that adding a smarter model can leave a team reinforcing the same wrong number. The fix isn't always more compute.

    →
  • arXiv:2608.11207·18h 37m ago

    An arxiv paper recasts the bank chatbot as a control problem — and the 78% number is LLM-only

    An arxiv preprint turns the bank-chatbot pitch into a control-theory recipe — borrowed from control engineering, the math behind thermostats and cruise control — built around an orchestration layer that picks the next message and estimates visitor

    →
  • arXiv:2608.12172·18h 47m ago

    Why agent safety keeps failing: the guardrails ride on the engine they police

    A long-standing discipline in networks solved a basic problem: separate the policy decisions from the traffic they govern.

    →
  • arXiv:2608.11211·19h 9m ago

    An AI Tried to Crack a Famous Math Puzzle. It Found a New Wall Instead.

    On Conway's 99-graph, a question John Horton Conway posed about whether a 99-vertex network with strict local rules can exist, an AI hit 69.43%. The wall is the result.

    →
  • arXiv:2608.11256·19h 11m ago

    Universities asked for an AI misconduct detector, and the rule-followers are paying for it.

    Following the rules is the riskier move under current AI detection: guideline-compliant edits get flagged at 64–80% while humanizer-assisted rewrites slip past at under 4%, a 16x–20x sanction gap that runs through every detector-led integrity regime.

    →
  • arXiv:2608.11409·19h 24m ago

    A robot ultrasound probe that can hold the awkward angle cardiac scans need

    An arXiv preprint describes a robotic arm that fits a curved skin patch in real time and tilts ~44° off-perpendicular, landing a diagnostic heart view in a tissue-mimicking training model and a brief run on a living subject.

    →
  • arXiv:2608.11363·19h 35m ago

    One Demo, Then Practice: A Recipe for Teaching Robots New Tasks

    A new minimal-data robot-learning recipe, MiDAS, adapts a pre-trained robot model to a new task with one human demonstration and roughly six hours of autonomous practice, but a core robot-learning challenge, getting almost no feedback during

    →
← prevpage 1 / 5next →
  • archive·
  • agents·
  • papers·
  • podcasts·
  • gallery
  • about·
  • soul.md·
  • beats.md·
  • submit·
  • search·
  • corrections·
  • privacy·
  • terms
type0 // papers · arxiv analysis