TLDRocket
Sign in
Latest An Anthropic AI model sent a false homicide tip to Philadelphia police — TechCrunch AI disqualification yields new Nikon Small World in Motion winner — Ars Technica McKinsey connects enterprise data through a knowledge graph for AI — SiliconANGLE Jeff Bezos says a 3-day workweek and more single-income households are... — Fortune Elon Musk just lost his trillionaire status again after three days—but... — Fortune Who’s taking your fast-food order? Chick-fil-A and McDonald’s have dif... — Fortune Ultra raises $62 million for fast-growing ‘robots as a service’ busine... — Fortune Nobody has been in charge for decades of the single most important par... — Fortune

The AI intelligence platform

Every AI story that matters — and the intelligence behind it.

TLDRocket reads all relevant sources, removes duplicate coverage, and publishes a short neutral summary of every story, linking back to the original. Free, no spam, unsubscribe anytime.

Add to Slack

Every story also updates live profiles event timelines weekly rankings the AI Market Index

Friday, 16 February 2024

Synthetic data: save money, time and carbon with open source

Hugging Face 2 years ago 46

Researchers demonstrated that using open-source language models to generate synthetic training data for custom sentiment analysis models costs 1,127 times less than using GPT-4 while maintaining equivalent accuracy. A custom RoBERTa model trained on synthetic data analyzed financial news for $2.7 and 0.12 kg CO2 compared to $3,061 and 735-1,100 kg CO2 with GPT-4, with 0.13-second latency versus multiple seconds. This approach enables companies to build task-specific models without the expense, latency, and third-party data exposure of commercial LLM APIs.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.