TLDRocket
Sign in
Latest Nebius looks to raise $4.5BN through bond issue — Tech.eu Also’s $3,500 e-bike is a $1 billion Trojan horse for autonomous trans... — Fortune Unitree, famous for its dancing robots, surges by 460% on its trading... — Fortune Exclusive: Replit taps OpenAI's low-cost Luna model for new 'Free Mode... — Fortune Adronite launches Codistry AI coding platform, claims half the token c... — SiliconANGLE Rundoo raises $30M to expand its AI-native operating system for small... — SiliconANGLE Temporal is in talks to raise $500M at a $12B pre-money valuation, mor... — Tech Funding News Etched raises $700M led by Jane Street, doubling to $21B and it still... — Tech Funding News

The AI intelligence platform

Every AI story that matters and the intelligence behind it.

TLDRocket reads all relevant sources, removes duplicate coverage, and publishes a short neutral summary of every story, linking back to the original. Free, no spam, unsubscribe anytime.

Add to Slack

Every story also updates live profiles event timelines weekly rankings the AI Market Index

Friday, 23 February 2024

🪆 Introduction to Matryoshka Embedding Models

Hugging Face 2 years ago 18

Matryoshka embedding models allow embeddings to be truncated to smaller dimensions while retaining performance, enabling storage and speed tradeoffs for tasks like retrieval and search. In experiments comparing a Matryoshka model to a standard model on STSBenchmark, the Matryoshka model preserved 98.37% of performance at 8.3% of full embedding size, versus 96.46% for the standard model. This approach makes it practical to deploy embedding systems across different storage budgets and processing speeds without significant accuracy loss.

Introducing the Red-Teaming Resistance Leaderboard

Hugging Face 2 years ago 40

Haize Labs released the Red-Teaming Resistance Leaderboard, a benchmark that tests language models against human-crafted adversarial prompts rather than algorithmically generated attacks that are unrealistic and easily detectable. The benchmark evaluates models across eight datasets (AdvBench, AART, Beavertails, Do Not Answer, RedEval-HarmfulQA, RedEval-DangerousQA, Student-Teacher Prompting, and SAP) and organizes harmful content into 14 specific violation categories including hate speech, fraud, and adult content. GPT-4 and Claude-2 lead the leaderboard with consistent robustness, while all tested models show greatest vulnerability to jailbreaks involving adult content, physical harm, and child harm.

Fine-Tuning Gemma Models in Hugging Face

Hugging Face 2 years ago 50 2 sources

Google Deepmind's Gemma language models are now available via Hugging Face in 2 billion and 7 billion parameter sizes, optimized for fine-tuning using parameter-efficient techniques on both GPUs and TPUs. The article demonstrates Low-Rank Adaptation (LoRA) fine-tuning, which reduces memory requirements by training only adapter layers rather than all model weights, enabling users to adapt Gemma models on platforms like Colab or Kaggle. Users can now customize Gemma's responses for specific tasks—such as formatting quote completion—without the computational cost of full model retraining.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.