TLDRocket
Sign in
Latest We're going to need default hard budget caps on pretty much everything — Simon Willison’s Weblog The Agent Said It Was Done. The Database Disagreed. — Hugging Face September sponsors-only newsletter — Simon Willison’s Weblog 9 insights from ‘Private Tech Trailblazers’: Vertical AI becomes the g... — SiliconANGLE OpenAI safety employee resigns, claiming the company’s ‘culture is bro... — TechCrunch Aleph Alpha’s Sovereign A.I. Model Kolibri Is No Match for the Open-We... — Trending Topics AI is speeding up exploits. Vulnerability spreadsheets can’t keep up. — The New Stack All the AI agents that can live in your text messages — TechCrunch

The AI intelligence platform

Every AI story that matters — and the intelligence behind it.

TLDRocket reads all relevant sources, removes duplicate coverage, and publishes a short neutral summary of every story, linking back to the original. Free, no spam, unsubscribe anytime.

Add to Slack

Every story also updates live profiles event timelines weekly rankings the AI Market Index

Thursday, 23 January 2025

Operator System Card

OpenAI 1 year ago 50

Operator is an AI system that uses OpenAI's safety frameworks to address risks including prompt injection attacks, jailbreaks, privacy violations, and security breaches. The system incorporates model-level protections, product mitigations, external red teaming assessments, and safety evaluations documented in a formal system card. OpenAI plans to continue refining these safeguards based on testing and deployment experience.

Mastering Long Contexts in LLMs with KVPress

Hugging Face 1 year ago 44

NVIDIA released KVPress, a toolkit that compresses the key-value cache used during language model inference to reduce memory consumption for long-context processing. Processing 1 million tokens with Llama 3-70B requires 327.6GB for the KV cache alone, but KVPress with a 50% compression ratio reduced peak memory usage from 45GB to 37GB on a 128k token context length. The compression enables faster decoding speeds, improving from 11 to 17 tokens per second on an A100 GPU, making long-context inference more practical for resource-constrained deployments.

SmolVLM Grows Smaller – Introducing the 256M & 500M Models!

Hugging Face 1 year ago 30 ● 2 sources

Hugging Face released SmolVLM-256M and SmolVLM-500M, vision language models with 256 million and 500 million parameters respectively. The 256M model is the smallest vision language model ever released and uses a 93-million-parameter vision encoder processing images at 4096 pixels per token, compared to 1820 pixels per token in the previous 2B version. These models enable vision language capabilities on consumer devices and low-cost data processing while maintaining performance on tasks like image captioning, document question-answering, and visual reasoning.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.