TLDRocket
Sign in
Latest Pay-per-inference for AI agents: How BlockRun and Incarna use Amazon B... — Amazon Web Services Ben Affleck is an AI nerd, and the internet is impressed — TechCrunch Popular AI leaderboard Arena nearly doubles valuation to $3.1B valuati... — TechCrunch OpenAI’s revenue is reportedly $20 billion less than previously projec... — TechCrunch Google brings agentic AI to Gemini, starting with businesses — TechCrunch Anthropic changes usage policy to ban model abuse and election interfe... — TechCrunch OpenAI’s math solutions aren’t meeting the field’s standards yet — TechCrunch 2026 Usage Policy update — Anthropic

The AI intelligence platform

Every AI story that matters — and the intelligence behind it.

TLDRocket reads all relevant sources, removes duplicate coverage, and publishes a short neutral summary of every story, linking back to the original. Free, no spam, unsubscribe anytime.

Add to Slack

Every story also updates live profiles event timelines weekly rankings the AI Market Index

Monday, 18 March 2024

Enterprise-ready trust and safety

OpenAI 2 years ago 49

Salesforce has integrated OpenAI's large language models into its platform to enable customer applications. The integration uses OpenAI's enterprise-grade models with built-in safety features and data handling designed for business use. Salesforce customers can now embed AI capabilities directly into their customer-facing applications without building separate AI infrastructure.

Reimagining the email experience with AI

OpenAI 2 years ago 42

Superhuman has partnered with OpenAI to integrate AI capabilities into its email platform. The service now offers AI-powered features including message summarization, smart replies, and automated categorization to streamline inbox management. Users can process email more quickly and reduce time spent on routine tasks through these automated functions.

Quanto: a PyTorch quantization backend for Optimum

Hugging Face 2 years ago 30

Hugging Face released Quanto, a PyTorch quantization backend for Optimum that reduces model size and computational costs by converting weights and activations to lower-precision data types like int8 or float8. The tool supports int2, int4, int8, and float8 weights across any model architecture and device (CPU, GPU, Apple Silicon), with accelerated int8-int8 and mixed-precision matrix multiplications on CUDA hardware. Quanto integrates directly into the transformers library, allowing developers to quantize models in a few lines of code without restricting themselves to specific model configurations or device types.

Easily Train Models with H100 GPUs on NVIDIA DGX Cloud

Hugging Face 2 years ago 44

Hugging Face launched Train on DGX Cloud, a service allowing Enterprise Hub organizations to fine-tune AI models using NVIDIA H100 GPUs through a no-code interface integrated into the Hugging Face Hub. The service charges $8.25 per GPU hour for H100 instances, with an example showing that fine-tuning Mistral 7B on 1,500 samples costs approximately $0.45. Users can now access GPU compute on-demand without writing training scripts, with fine-tuned models automatically saved to private repositories. (Note: The service was deprecated as of April 10, 2025.)

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.