Blog
Updates and guides from us, plus hand-picked reads from tech & AI blogs.
Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps
Training a coding model to paint watercolours with TRL and OpenEnv
Real-Time Intelligence with IBM Time Series Models on Confluent
Reach audiences
Advertise in front of researchers, engineers, and readers.
TimesFM-3: A zero-shot foundation model for multivariate forecasting
Data Management
GigaPath-Flash and GigaTIME-Flash: Toward population-scale discovery with efficient pathology foundation models
What if pathology foundation models could do more with less? GigaPath-Flash and GigaTIME-Flash cut computational demands while maintaining strong performance, opening the door to larger studies and broader exploration. The post GigaPath-Flash and GigaTIME-Flash: Toward population-scale discovery with efficient pathology foundation models appeared first on Microsoft Research.
Planetary prediction engine: Automating global models via Earth AI
Earth AI
GlucoFM: Foundation model for continuous glucose monitoring
Health & Bioscience
Training and Finetuning Multi-Vector Embedding Models with Sentence Transformers
We’re on a journey to advance and democratize artificial intelligence through open source and open science.
Quantization-Aware Healing: a compressed, 4-bit model that outperforms its full-precision original
A Blog post by Multiverse Computing on Hugging Face
How mobility gives language models a deeper understanding of place
Algorithms & Theory
When AI art has no author: Study finds generated images often can’t be traced to training data
A new method for surgically removing training examples from a model reveals that as datasets grow, the link between what a model learns and what it produces dissolves.