TLDRocket
Sign in
Latest Exa Launches Agent Ultra: A Subagent Swarm Deep Research API Built for... — MarkTechPost End-to-End Multimodal Data Augmentation and Adversarial Robustness Ben... — MarkTechPost Government Hacks and “Agent Spam”: OpenAI Reports Dozens More Rogue AI... — Trending Topics At Meta Connect, the company’s smart glasses were everywhere — TechCrunch What to expect at Dell’s AI Leadership Symposium: Join theCUBE Sept. 2... — SiliconANGLE OpenRouter: from Seed to Stripe — with OpenRouter’s Alex Atallah & AMP... — Latent Space Liquid AI Releases LFM2.5-VL-3B-DSpark: Speculative Decoding for Visio... — MarkTechPost Crusoe abandons $1.25B plan to use Boom turbines at AI data centers — TechCrunch

The AI intelligence platform

Every AI story that matters — and the intelligence behind it.

TLDRocket reads all relevant sources, removes duplicate coverage, and publishes a short neutral summary of every story, linking back to the original. Free, no spam, unsubscribe anytime.

Add to Slack

Every story also updates live profiles event timelines weekly rankings the AI Market Index

Thursday, 18 June 2026

MosaicLeaks: Can your research agent keep a secret?

Hugging Face 3 months ago 42

Research agents that combine private documents with web searches leak sensitive information through the cumulative pattern of their queries, even when individual searches appear innocuous. Across tested models, answer or full-information leakage reached 34.0% in baseline agents and climbed to 51.7% when agents were trained only for task performance. The Privacy-Aware Deep Research training method reduced leakage to 9.9% while maintaining 58.7% strict chain success by rewarding agents for constructing queries that avoid revealing private details.

Deep Learning Weekly: Issue 460

Deep Learning Weekly 3 months ago 12

Deep Learning Weekly issue 460 covers multiple AI model releases and research advances, including Moonshot AI's Kimi K2.7 Code with 1T parameters achieving 31.5% benchmark improvements, Stanford's DeLM reducing multi-agent task costs by 50%, and Z.ai's GLM-5.2 with 753B parameters and 1M-token context window. Research papers include Data Journalist Agent for automated multimodal news generation and FastContext for specialized repository exploration in coding agents, which reduces token consumption by up to 60% while improving task resolution rates by 5.5%. The issue also features tools for LLM observability, cost optimization techniques for Claude Code, and infrastructure developments in multimodal model serving.

Launch HN: TesterArmy (YC P26) – Agents that test web and mobile apps

tester.army 3 months ago 39

TesterArmy, a Y Combinator-backed startup, launched an AI-agent-based testing platform that uses natural language to define and execute end-to-end tests for web and mobile apps. The platform has grown to 30+ daily users over several months and has caught production bugs including timezone issues, API failures, and checkout regressions that human testers missed. Teams can now skip manual test maintenance and instead let agents handle test creation, execution, and scheduling via CLI integration.

The first big exit in AI

Ben's Bites 3 months ago 12 ● 3 sources

SpaceX is acquiring Cursor, an AI coding assistant, for $60 billion in stock. The acquisition price reflects the significant valuation of AI development tools in the current market. The deal marks a major consolidation in the AI tooling space, with Cursor's technology becoming part of SpaceX's broader operations.

Sound Waves Give Neuromorphic Chips a Brain-Simulating Edge

IEEE Spectrum 3 months ago 21

Researchers developed an acoustic synapse using sound waves and phase bits to create neuromorphic hardware that mimics biological neurons more closely than electronic alternatives. The device achieved 96.7 percent accuracy on iris flower classification using 39 parameters and 20 percent faster than conventional neural networks while consuming one-tenth the power of current electronic neuromorphic hardware. This approach enables more compact and energy-efficient neuromorphic systems capable of mimicking neuromodulation and handling complex pattern recognition tasks with simpler designs.

Anthropic ships major Claude Design overhaul with design system imports, code round-trips, and a fix for its token-burning problem

VentureBeat 3 months ago 21

Anthropic released an updated Claude Design with reduced token usage, new capabilities to import design systems from multiple sources, and code round-trip functionality for developers. The tool now supports importing design systems from GitHub, design files, or direct uploads to help teams maintain consistent branding. Teams can integrate design-to-code workflows more efficiently, lowering operational costs for design at scale.

Mobileye is entering the US robotaxi market with standalone service

Ars Technica 3 months ago 9

Mobileye announced it will operate its own robotaxi service in the US in addition to supplying autonomous driving technology to automakers and mobility providers. The company plans to deploy 100 robotaxis initially and scale to around 17,000 within five years if the pilot succeeds. This direct operation allows Mobileye to gain real-world experience while maintaining existing partnerships with automakers like Volkswagen and Lyft.

Uneven Frontiers

X 3 months ago 20

AI is expected to accelerate and reduce costs in certain areas of drug development, but clinical trials will remain a significant bottleneck because recruiting patients and collecting long-term safety data cannot be substantially sped up by artificial intelligence. The physical and regulatory constraints of human testing create a ceiling on how much AI can improve the overall timeline and cost of bringing new drugs to market. Consequently, AI's impact on pharma will be uneven, transforming discovery and early research while leaving late-stage development largely unchanged.

Anthropic Employees Accuse Trump Administration of Targeting Them

The New York Times 3 months ago 41

Anthropic employees and over 150 cybersecurity experts have signed a letter criticizing the Trump administration's ban on Anthropic's AI models, calling the restrictions unfair and urging their removal. The letter represents the first significant public backlash to the ban from security professionals in the field. The criticism may influence whether the administration reconsiders or modifies its restrictions on the AI company's operations.

Improving health intelligence in ChatGPT

OpenAI 3 months ago 44

GPT-5.5 Instant has been integrated into ChatGPT to enhance its health and wellness responses through improved reasoning and contextual understanding. The model incorporates physician-informed evaluations to provide more clinically grounded guidance. This update allows ChatGPT to deliver clearer health information while maintaining appropriate disclaimers about medical advice limitations.

The Sequence Opinion #879: When Tokens Become Balance Sheet Items

Substack 3 months ago 25

Companies are increasingly treating AI tokens as formal accounting line items and balance sheet expenses as tokens become the fundamental unit of the AI economy. Major enterprises now forecast and report token consumption similar to traditional accounting metrics, though standard measurement frameworks remain underdeveloped. This shift may drive demand for new software infrastructure and tools to manage token economics across organizations.

Blaming China for Datacenter NIMBYism Is Cope

ChinaTalk 3 months ago 4

The article argues that while China has conducted influence operations against U.S. datacenters, blaming foreign propaganda for community opposition is misleading and ignores legitimate local concerns that can be addressed through financial incentives. Major AI companies like Microsoft, Google, Amazon, and Meta plan to spend $725 billion on capex in 2026, with $500 billion domestically, giving them substantial resources to compensate communities. The author proposes datacenter operators could resolve resistance by offering annual payments to residents (such as $10,000 per person for 3.8% of revenue), investing in local schools, or building community amenities like pickleball courts, rather than the token $10 million commitment OpenAI announced for Stargate.

How Domyn and AISquared built on Ai2's open releases

Allen Institute (AI2) 3 months ago 23

Domyn and AISquared, two AI labs serving regulated industries, built commercial models by fine-tuning Ai2's open-source releases: AISquared created Bolt from Olmo for enterprise workflows, while Domyn built a reasoning model using Dolma and Dolci datasets. AISquared's Bolt Instruct comes in sizes of 1B, 7B, and 32B parameters, with customers reporting roughly 50% reductions in infrastructure costs. The full transparency and documented provenance of Ai2's open artifacts—including training data, code, and architectural details—enabled both companies to meet compliance requirements for regulated sectors like finance and government, making open-source alternatives viable for customers with strict auditability constraints.

Using AI to help physicians diagnose rare genetic diseases affecting children

OpenAI 3 months ago 40

Researchers deployed an AI reasoning model to assist physicians in diagnosing rare genetic diseases in children, successfully identifying new cases that had previously gone undiagnosed. The system produced 18 new diagnoses from cases where standard medical approaches had failed to reach a conclusion. This capability could reduce diagnostic delays for families seeking answers about their children's genetic conditions.

Is it agentic enough? Benchmarking open models on your own tooling

Hugging Face 3 months ago 6

Researchers benchmarked how well different language models can use the transformers library by measuring not just whether they got the right answer, but how much effort it took them to get there. The evaluation framework tested each task across three different tool access levels (bare install, cloned source code, or packaged skill) and found that adding a CLI reduced median task completion time but increased token consumption by 60% due to agents reading documentation. The results show that library design significantly affects agent efficiency—smaller models benefited more from curated documentation while larger models sometimes performed better with full source code access.

Beyond LoRA: Can you beat the most popular fine-tuning technique?

Hugging Face 3 months ago 40

Hugging Face benchmarked over 40 parameter-efficient fine-tuning techniques to compare their performance against LoRA, which dominates 98.4% of fine-tuning implementations on their hub. On mathematical reasoning tasks, LoRA achieved 53.2% accuracy using 22.6 GB of memory, while on image generation, the OFT technique scored 0.708 similarity versus LoRA's 0.697 with lower memory (9.01 GB vs 9.97 GB). Users should evaluate multiple PEFT techniques on their own datasets rather than defaulting to LoRA, as different methods offer better tradeoffs depending on whether accuracy or memory efficiency is prioritized.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.