TLDRocket
Sign in
Latest Nebius looks to raise $4.5BN through bond issue — Tech.eu Also’s $3,500 e-bike is a $1 billion Trojan horse for autonomous trans... — Fortune Unitree, famous for its dancing robots, surges by 460% on its trading... — Fortune Exclusive: Replit taps OpenAI's low-cost Luna model for new 'Free Mode... — Fortune Adronite launches Codistry AI coding platform, claims half the token c... — SiliconANGLE Rundoo raises $30M to expand its AI-native operating system for small... — SiliconANGLE Temporal is in talks to raise $500M at a $12B pre-money valuation, mor... — Tech Funding News Etched raises $700M led by Jane Street, doubling to $21B and it still... — Tech Funding News

The AI intelligence platform

Every AI story that matters and the intelligence behind it.

TLDRocket reads all relevant sources, removes duplicate coverage, and publishes a short neutral summary of every story, linking back to the original. Free, no spam, unsubscribe anytime.

Add to Slack

Every story also updates live profiles event timelines weekly rankings the AI Market Index

Thursday, 1 February 2024

Patch Time Series Transformer in Hugging Face

Hugging Face 2 years ago 35

PatchTST, a time series forecasting model based on Transformers, was added to Hugging Face with documentation showing how to train and apply it to electricity data using transfer learning. The model divides time series into patches of configurable length (the example uses patch_length=16 with context_length=512) and reduces computational complexity quadratically compared to processing full sequences. Users can now train PatchTST directly on datasets, perform zero-shot forecasting on new domains, and fine-tune pretrained models using the Hugging Face Trainer API.

Hugging Face Text Generation Inference available for AWS Inferentia2

Hugging Face 2 years ago 11

Hugging Face Text Generation Inference became generally available on AWS Inferentia2 through Amazon SageMaker for deploying large language models in production. The solution supports popular models like Llama and Mistral, with pre-compiled configurations cached for batch size 2-4 and sequence length 2048 to avoid the 45-minute compilation process. Customers can now deploy LLMs on Inferentia2 as a cost-effective alternative to GPUs, with deployment taking 10-15 minutes on ml.inf2.8xlarge instances.

Constitutional AI with Open LLMs

Hugging Face 2 years ago 41

Researchers released Constitutional AI tools and datasets enabling open-source language models to self-align by critiquing their own outputs against user-defined principles. The team published the llm-swarm tool for generating synthetic training data at scale on GPU clusters, along with aligned Mistral 7B models and datasets based on both Anthropic's and a Grok-inspired constitution. This allows practitioners to customize model behavior without collecting expensive human feedback by having models identify and revise responses that violate constitutional principles.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.