TLDRocket
Sign in
Latest Nebius looks to raise $4.5BN through bond issue — Tech.eu Also’s $3,500 e-bike is a $1 billion Trojan horse for autonomous trans... — Fortune Unitree, famous for its dancing robots, surges by 460% on its trading... — Fortune Exclusive: Replit taps OpenAI's low-cost Luna model for new 'Free Mode... — Fortune Adronite launches Codistry AI coding platform, claims half the token c... — SiliconANGLE Rundoo raises $30M to expand its AI-native operating system for small... — SiliconANGLE Temporal is in talks to raise $500M at a $12B pre-money valuation, mor... — Tech Funding News Etched raises $700M led by Jane Street, doubling to $21B and it still... — Tech Funding News

The AI intelligence platform

Every AI story that matters and the intelligence behind it.

TLDRocket reads all relevant sources, removes duplicate coverage, and publishes a short neutral summary of every story, linking back to the original. Free, no spam, unsubscribe anytime.

Add to Slack

Every story also updates live profiles event timelines weekly rankings the AI Market Index

Tuesday, 7 November 2023

Comparing the Performance of LLMs: A Deep Dive into Roberta, Llama 2, and Mistral for Disaster Tweets Analysis with Lora

Hugging Face 2 years ago 36

Researchers compared three language models—RoBERTa, Llama 2, and Mistral 7B—fine-tuned with LoRA (Low-Rank Adaptation) for classifying disaster-related tweets. The study used 7,613 training tweets split into train (6,090), validation (1,523), and test (3,263) samples, with class weights of 1.16 for positive and 0.88 for negative examples to address imbalance. Fine-tuning was performed with LoRA to reduce trainable parameters while maintaining performance on the sequence classification task.

Introducing Prodigy-HF: a direct integration with Hugging Face

Hugging Face 2 years ago 3

Explosion released Prodigy-HF, a plugin that integrates its Prodigy annotation tool directly with Hugging Face models and infrastructure. Users can now fine-tune transformer models like distilbert-base-uncased on annotated data with a single command and upload datasets to the Hugging Face Hub via the hf.upload recipe. This enables faster iteration on domain-specific NLP tasks by allowing models trained on annotated data to be reused for further annotation work.

Make your llama generation time fly with AWS Inferentia2

Hugging Face 2 years ago 5

AWS and Hugging Face enabled text generation with Llama 2 models on AWS Inferentia2 accelerators using the optimum-neuron library, which compiles and deploys large language models to specialized hardware. The Llama 2 7B model achieves encoding times of 0.5 seconds for 256 input tokens and throughput of 227–750 tokens per second depending on configuration, while the 13B model reaches 145–504 tokens per second. Users can now export models from Hugging Face, compile them for Inferentia2 with static shape constraints, and generate text using standard transformer APIs or simplified pipeline wrappers.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.