One Useful Thing
·
1 month ago
● 3 sources
The article examines how AI tools can undermine human cognitive development when used as shortcuts but can enhance learning when used intentionally, citing research showing students using ChatGPT for homework performed worse on tests while those guided by AI tutors improved by 0.15 standard deviations. The author argues that the design of AI systems prioritizing frictionless use encourages cognitive surrender, similar to how 758 consultants at Boston Consulting Group using GPT-4 failed to catch errors the AI made on unfamiliar problems. The piece urges conscious choice about which cognitive tasks to delegate to AI versus keep for human development, warning that defaults being set now by AI companies and institutions may be difficult to reverse.
Interconnects
·
1 month ago
Open-weight models have not yet reached the agentic capability level demonstrated by Claude's Opus 4.5 in December 2025, with the gap likely persisting for 12+ months despite steady benchmark improvements. Google's Gemini 3.5 Flash lacks a meaningful competitor to Claude Code and Codex, suggesting Chinese labs face similar resource constraints and will specialize in automated enterprise agents rather than competing directly with U.S. frontier models. Existing power structures including governments and the Pope are beginning to assert control over AI development, while social and political resistance to data center expansion in the U.S. is creating potential conflict over who shapes the technology's future.
Amazon Science
·
1 month ago
Researchers introduced a method for training large language models on multiple diverse reasoning paths for the same problem, using global forking tokens and set-supervised fine-tuning to prevent mode collapse where different tokens produce identical outputs. Performance improved 5% to 7% on standard benchmarks (with gains of 6.84% on AIME 2025 Pass@1 compared to baseline methods), demonstrating that enabling models to learn distinct reasoning strategies and select the appropriate one per question directly improves accuracy. This approach enables LLMs to solve problems through multiple qualitatively different strategies like algebraic manipulation or geometric reasoning, rather than collapsing to a single solution method.
Menlo Ventures
·
1 month ago
OpenRouter, an AI model routing platform, raised $113M in Series B funding at a $1.3B valuation after growing to 8M developers. The company now processes approximately 1.5 quadrillion tokens annually, representing 15-30% of Google's token run rate and 20-40% of OpenAI's run rate. OpenRouter enables developers and enterprises to access 400+ AI models through a single interface with unified billing, improved reliability, and cost optimization features.
Ben's Bites
·
1 month ago
A tech newsletter discusses whether SaaS is declining as users increasingly build custom solutions with AI agents and modular tools instead of monolithic platforms. Perplexity open-sourced Bumblebee, a safety scanner for developer machines, and WorkOS released auth.md, an open protocol for agents to register for web services. The shift toward API-first and composable software tools is reshaping how users expect to customize and control their software features.
Import AI
·
1 month ago
Dario Amodei, Anthropic's cofounder, argues that rapid AI progress requires societies to actively shape the technology's future rather than passively react to it. He cites AI achieving gold in the International Math Olympiad in July 2025 and co-authoring mathematical proofs the same year as evidence of accelerating capabilities, with self-improving systems potentially arriving within two years. The stakes demand immediate decisions about how to distribute AI's benefits and manage existential risks, since continued development under geopolitical competition makes coordinated global slowdowns unlikely.
Last Week in AI
·
1 month ago
● 2 sources
Google unveiled Gemini 3.5 and the Gemini Spark agent at I/O 2026, alongside multimodal video generation capabilities and research tools, while Elon Musk lost his OpenAI lawsuit on statute-of-limitations grounds and Anthropic secured a $30 billion funding round at a $900 billion valuation. Cerebras' IPO surged 90%, coding-agent competition accelerated with Cursor Composer 2.5 and xAI's Grok Build, and OpenAI resolved an 80-year-old Erdős geometry problem. AI research advanced in interpretability, cyber capabilities, and autonomous hacking, while policy efforts addressed deepfakes and image provenance.