TLDRocket
Sign in
Latest Nebius looks to raise $4.5BN through bond issue — Tech.eu Also’s $3,500 e-bike is a $1 billion Trojan horse for autonomous trans... — Fortune Unitree, famous for its dancing robots, surges by 460% on its trading... — Fortune Exclusive: Replit taps OpenAI's low-cost Luna model for new 'Free Mode... — Fortune Adronite launches Codistry AI coding platform, claims half the token c... — SiliconANGLE Rundoo raises $30M to expand its AI-native operating system for small... — SiliconANGLE Temporal is in talks to raise $500M at a $12B pre-money valuation, mor... — Tech Funding News Etched raises $700M led by Jane Street, doubling to $21B and it still... — Tech Funding News

The AI intelligence platform

Every AI story that matters and the intelligence behind it.

TLDRocket reads all relevant sources, removes duplicate coverage, and publishes a short neutral summary of every story, linking back to the original. Free, no spam, unsubscribe anytime.

Add to Slack

Every story also updates live profiles event timelines weekly rankings the AI Market Index

Tuesday, 2 April 2024

Qwen1.5-32B: Fitting the Capstone of the Qwen1.5 Language Model Series

GitHub Pages 2 years ago 23

Alibaba released Qwen1.5-32B, a 32-billion-parameter open-source language model designed to balance performance with efficiency and lower memory requirements. The model aims to address the resource constraints of larger models like Qwen1.5-72B by fitting the approximately 30 billion parameter range identified by the community as optimal for practical deployment. This addition to the Qwen1.5 series enables developers to use a capable model with reduced computational costs and faster inference speeds compared to larger alternatives.

Bringing serverless GPU inference to Hugging Face users

Hugging Face 2 years ago 3

Hugging Face and Cloudflare launched an integration enabling developers to run open-source AI models as serverless APIs on Cloudflare's GPU infrastructure without managing their own servers. An example RAG application handling 1,000 requests daily with Llama 2 7B would cost approximately $1 per day under the pay-per-request pricing model. Developers can now deploy popular models like Llama, Gemma, and Mistral directly from Hugging Face's Hub using either Cloudflare's REST API or AI SDK. Note: The article's November 2024 update states this integration is no longer available and directs users to alternative deployment options.

Customizing models for legal professionals

OpenAI 2 years ago 43

Harvey, a legal AI platform, partnered with OpenAI to develop a custom-trained model designed specifically for legal professionals. The model was built by fine-tuning OpenAI's technology using proprietary legal datasets and workflows that Harvey had developed. This allows Harvey to offer legal practitioners a tool optimized for their domain-specific needs rather than using general-purpose AI.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.