The Orthrus framework achieves up to 7.8× tokens per forward pass on Qwen3 models while maintaining a provably identical output distribution to the original. Its dual-view architecture shares a single KV cache between autoregressive and diffusion pathways.
#open-source
RSS FeedThe popular text-generation-webui project, rebranded as TextGen, has relaunched as a no-install native desktop app for Windows, Linux, and macOS. Built on a minimal Electron integration, it positions itself as a fully open-source alternative to LM Studio.
The UK-based AI startup founded by former OpenAI, DeepMind, and Meta researchers raised $650M at a $4.65B valuation to build recursively self-improving AI systems, backed by NVIDIA and GV.
The RPCS3 PS3 emulator's GitHub has been overwhelmed by low-quality AI-generated pull requests. Developers issued a public warning, threatening bans for those who submit AI code without disclosure.
Carnegie Mellon University and the Bosch Center for AI developed HTD (Humanoid Transformer with Touch Dreaming), which trains robots to anticipate future tactile signals. Across five dexterous real-world tasks, HTD achieved 90.9% higher success rates over vision-only baselines.
Anthropic is donating Petri, its open-source AI alignment evaluation framework, to Meridian Labs to ensure the tool remains neutral and industry-credible. Petri 3.0 brings major improvements in adaptability, realism, and depth.
The Allen Institute for AI released MolmoAct 2 on May 5, a fully open-source 7B-parameter robot foundation model that surpasses Physical Intelligence's π0.5 across simulation and real-world tasks, achieving up to 87.1% success on zero-shot manipulation.
Chinese AI startup Moonshot AI raised $2 billion led by Meituan at a $20 billion valuation, bringing total capital to $3.9 billion over six months — making it China's most heavily funded LLM startup.
Engineer Aaed Musa has released CARA 2.0, an improved open-source quadruped robot dog featuring better kinematics, joint design, and real-time control over the original version.
Sakana AI released KAME, a tandem speech-to-speech architecture that pairs a low-latency front-end S2S model with a back-end LLM via an oracle stream, achieving MT-Bench 6.43 with near-zero response latency and eliminating the typical 2.1-second pipeline delay.
Poolside AI released Laguna XS.2 on April 28, 2026 under Apache 2.0 — a 33B total/3B active MoE model purpose-built for agentic coding, scoring 68.2% on SWE-bench Verified and deployable on a single consumer GPU.
Released April 29, 2026 under Modified MIT license, Mistral Medium 3.5 consolidates the company's chat, reasoning, and coding models into one 128B dense open-weight model with 256K context, scoring 77.6% on SWE-bench Verified.