deployedbyai

The Deploy Log

Hugging Face

Every deployment from Hugging Face in The Deploy Log, 74 rows, newest first.

WebGPU kernels

Hugging Face released @huggingface/kernels with 207 WebGPU kernels, 2.57x faster than ORT WebGPU by geometric mean across 809 matching test cases on an Apple M4 GPU.

Hugging Face | Models and APIs | In Edition 50, Saturday 5 Sep 2026

SOURCE Hugging Face, Introducing @huggingface/kernels: 200+ WebGPU Kernels for Local AI

Muse Glimmer 30B

Meta released Muse Glimmer, a 30B multimodal model under Apache 2.0, with day-0 support in transformers, llama.cpp, vLLM, and Inference Endpoints.

Hugging Face | Models and APIs | Published Monday 10 Aug 2026

SOURCE Hugging Face, Meta is back with Muse Glimmer: local, agentic, multimodal, and open source!

Baseten Inference Provider

Baseten is now a supported Inference Provider on the Hugging Face Hub, launching support for conversational and text-generation tasks with open-weight LLMs such as Kimi K3, DeepSeek V4 Flash, and GLM-5.2.

Hugging Face | Models and APIs | In Edition 46, Saturday 8 Aug 2026

SOURCE Hugging Face, Baseten on Hugging Face Inference Providers

Nunchaku Lite in Diffusers

Nunchaku Lite brings 4-bit diffusion inference to Diffusers, cutting peak VRAM by up to 50% and improving latency by 30%.

Hugging Face | Models and APIs | Published Thursday 23 Jul 2026

SOURCE Hugging Face, Bringing Nunchaku 4-bit Diffusion Inference to Diffusers

LeRobot v0.6.0

LeRobot v0.6.0 ships world model policies VLA-JEPA, FastWAM, and LingBot-VA, new VLAs including GR00T N1.7 and MolmoAct2, reward models Robometer and TOPReward, and six new simulation benchmarks under lerobot-eval.

Hugging Face | Breakthroughs | In Edition 42, Saturday 11 Jul 2026

SOURCE Hugging Face, LeRobot v0.6.0: Imagine, Evaluate, Improve

vLLM transformers backend

The transformers vLLM backend now meets or beats native throughput on Qwen3 4B dense, 32B dense, and 235B-parameter FP8 MoE models, using torch.fx static analysis and ast source rewriting to apply inference-specific layer fusions at runtime.

Hugging Face | Models and APIs | In Edition 42, Saturday 11 Jul 2026

SOURCE Hugging Face, Native-speed vLLM transformers modeling backend

Gemma 4 voice AI

Hugging Face and Cerebras demonstrate a real-time speech-to-speech pipeline pairing Gemma 4 31B on Cerebras with Nvidia's Parakeet for speech recognition and Alibaba's Qwen3TTS for text-to-speech, already powering more than 9,000 Reachy Mini robots.

Hugging Face | Models and APIs | In Edition 41, Saturday 4 Jul 2026

SOURCE Hugging Face, Hugging Face and Cerebras bring Gemma 4 to real-time voice AI

PP-OCRv6

PP-OCRv6 scales from 1.5M to 34.5M parameters across tiny, small, and medium tiers, with the medium and small tiers supporting 50 languages and the medium tier reaching 86.2% detection Hmean and 83.2% recognition accuracy.

Hugging Face | Models and APIs | In Edition 40, Saturday 27 Jun 2026

SOURCE Hugging Face, PP-OCRv6 on Hugging Face: 50-Language OCR from 1.5M to 34.5M Parameters

GLM-5.2

GLM-5.2 is an MIT-licensed open-source model with a 1M-token context. It scores 81.0 on Terminal-Bench 2.1 and 62.1 on SWE-bench Pro, and is the highest-ranked open-source model on FrontierSWE, PostTrainBench, and SWE-Marathon.

Hugging Face | In Edition 39, Saturday 20 Jun 2026

SOURCE Hugging Face, GLM-5.2: Built for Long-Horizon Tasks

Strands Robots

Strands Robots is an Apache 2.0 SDK from AWS that exposes robot abstractions, simulation, and the LeRobot stack as AgentTools, with sim and hardware datasets sharing the same on-disk format.

Hugging Face | Robotics, hardware and chips | In Edition 39, Saturday 20 Jun 2026

SOURCE Hugging Face, From the Hugging Face Hub to robot hardware with Strands Agents and LeRobot

Mellum2

JetBrains released Mellum2, a 12B-parameter Mixture-of-Experts model that activates only 2.5B parameters per token and delivers more than 2x faster inference than similar-sized models under the Apache 2.0 license.

Hugging Face | Models and APIs | In Edition 37, Saturday 6 Jun 2026

SOURCE Hugging Face, Introducing Mellum2: A 12B Mixture-of-Experts Model by JetBrains

Nemotron 3.5 Content Safety

NVIDIA released Nemotron 3.5 Content Safety, a 4B-parameter model that unifies multimodal input, 12-language explicit coverage, custom policy enforcement, and auditable reasoning traces in one inference call.

Hugging Face | Models and APIs | In Edition 37, Saturday 6 Jun 2026

SOURCE Hugging Face, Nemotron 3.5 Content Safety: Customizable Multimodal Safety for Global Enterprise AI

OlmoEarth v1.1

OlmoEarth v1.1 is a new family of Earth observation models that cuts compute costs by up to 3x while maintaining OlmoEarth v1's performance on research benchmarks.

Hugging Face | Models and APIs | Published Tuesday 19 May 2026

SOURCE Hugging Face, OlmoEarth v1.1: A more efficient family of Earth observation models

PaddleOCR 3.5

PaddleOCR 3.5 brings OCR and document parsing tasks closer to the Hugging Face ecosystem, with supported models able to run with Transformers as an inference backend.

Hugging Face | Models and APIs | Published Monday 18 May 2026

SOURCE Hugging Face, PaddleOCR 3.5: Running OCR and Document Parsing Tasks with a Transformers Backend

Granite Embedding Multilingual R2

IBM released two Apache 2.0 multilingual embedding models built on ModernBERT: a 97M-parameter compact model scoring 60.3 on MTEB Multilingual Retrieval and a 311M full-size model scoring 65.2, both covering 200+ languages with 32K-token context.

Hugging Face | Models and APIs | In Edition 34, Saturday 16 May 2026

SOURCE Hugging Face, Granite Embedding Multilingual R2: Open Apache 2.0 Multilingual Embeddings with 32K Context

Nemotron 3 Nano Omni

NVIDIA released Nemotron 3 Nano Omni, a 30B-A3B omni-modal model for documents, audio, video, and agentic computer use, with OCRBenchV2-En 65.8, MMLongBench-Doc 57.5, Video-MME 72.2, and VoiceBench 89.4, and checkpoints in BF16, FP8, and NVFP4 on Hugging Face.

Hugging Face | Models and APIs | In Edition 32, Saturday 2 May 2026

SOURCE Hugging Face, Introducing NVIDIA Nemotron 3 Nano Omni

DeepInfra on HF

DeepInfra is now a supported Inference Provider on Hugging Face Hub, offering over 100 models including DeepSeek V4, Kimi-K2.6, GLM-5.1, with PRO users getting $2 monthly Inference credits.

Hugging Face | Models and APIs | Published Wednesday 29 Apr 2026

SOURCE Hugging Face, DeepInfra on Hugging Face Inference Providers

Granite 4.1 LLMs

IBM released Granite 4.1, a family of dense LLMs (3B, 8B, 30B) trained on ~15T tokens with 512K context, Apache 2.0, with the 8B matching the previous 32B MoE.

Hugging Face | Models and APIs | Published Wednesday 29 Apr 2026

SOURCE Hugging Face, Granite 4.1 LLMs: How They’re Built

DeepSeek-V4

DeepSeek released V4 today with two MoE checkpoints on the Hub: DeepSeek-V4-Pro at 1.6T total parameters with 49B active, and DeepSeek-V4-Flash at 284B total with 13B active, both with a 1M-token context window.

Hugging Face | Models and APIs | In Edition 31, Saturday 25 Apr 2026

SOURCE Hugging Face, DeepSeek-V4: a million-token context that agents can actually use

Transformers.js Chrome extension

Hugging Face published a guide on using Transformers.js in a Chrome extension, with a demo powered by Gemma 4 E2B for local AI features.

Hugging Face | Tools and products | Published Thursday 23 Apr 2026

SOURCE Hugging Face, How to Use Transformers.js in a Chrome Extension

Multimodal Sentence Transformers

Sentence Transformers v5.4 adds multimodal embedding and reranker models that encode and compare text, images, audio, and video through the same API, with Qwen3-VL-Embedding-2B and Qwen3-VL-Reranker-2B supported.

Hugging Face | Models and APIs | In Edition 29, Saturday 11 Apr 2026

SOURCE Hugging Face, Multimodal Embedding & Reranker Models with Sentence Transformers

Waypoint-1.5

Overworld released Waypoint-1.5, a real-time video world model with 720p and 360p tiers that runs locally on RTX 3090 through 5090 hardware at up to 60 FPS.

Hugging Face | In Edition 29, Saturday 11 Apr 2026

SOURCE Hugging Face, Waypoint-1.5: Higher-Fidelity Interactive Worlds for Everyday GPUs

Falcon Perception

Falcon Perception is a 0.6B-parameter early-fusion Transformer that reaches 68.0 Macro-F1 on SA-Co versus 62.3 for SAM 3, and Falcon OCR is a 0.3B model scoring 80.3 on olmOCR and 88.6 on OmniDocBench.

Hugging Face | Models and APIs | In Edition 28, Saturday 4 Apr 2026

SOURCE Hugging Face, Falcon Perception

Granite 4.0 3B Vision

Granite 4.0 3B Vision is available on Hugging Face under Apache 2.0 and leads on PubTablesV2 cropped at 92.1 and full-page at 79.3, OmniDocBench at 64.0, and TableVQA at 88.1.

Hugging Face | Models and APIs | In Edition 28, Saturday 4 Apr 2026

SOURCE Hugging Face, Granite 4.0 3B Vision: Compact Multimodal Intelligence for Enterprise Documents

TRL v1.0

TRL v1.0 is released with a stable core following semantic versioning and an experimental layer, and the library is downloaded 3 million times a month.

Hugging Face | Models and APIs | In Edition 28, Saturday 4 Apr 2026

SOURCE Hugging Face, TRL v1.0: Post-Training Library Built to Move with the Field

mRNA models 25 species

OpenMed trained 4 production mRNA language models across 25 species in 55 GPU-hours for $165, with CodonRoBERTa-large-v2 reaching perplexity 4.10 and Spearman CAI correlation 0.40.

Hugging Face | Models and APIs | In Edition 28, Saturday 4 Apr 2026

SOURCE Hugging Face, Training mRNA Language Models Across 25 Species for $165

gradio.Server

Gradio released gradio.Server, letting developers pair any custom frontend with Gradio's backend, including queuing and ZeroGPU support.

Hugging Face | Models and APIs | Published Wednesday 1 Apr 2026

SOURCE Hugging Face, gradio.Server: Any Custom Frontend with Gradio's Backend

EVA framework

ServiceNow AI released EVA, an end-to-end evaluation framework for voice agents that jointly scores accuracy (EVA-A) and experience (EVA-X), with an airline dataset of 50 scenarios and results for 20 systems.

Hugging Face | Models and APIs | Published Tuesday 24 Mar 2026

SOURCE Hugging Face, A New Framework for Evaluating Voice Agents (EVA)

Holotron-12B

Holotron-12B, post-trained from NVIDIA Nemotron-Nano-2 VL, raises WebVoyager from 35.1% to 80.5% and achieves over 2x higher throughput than Holo2-8B on a single H100.

Hugging Face | Models and APIs | In Edition 26, Saturday 21 Mar 2026

SOURCE Hugging Face, Holotron-12B - High Throughput Computer Use Agent

LeRobot v0.5.0

LeRobot v0.5.0 adds full Unitree G1 humanoid support, Pi0-FAST autoregressive VLAs, Real-Time Chunking, streaming video encoding with zero wait between episodes, and EnvHub for loading simulation environments from the Hub.

Hugging Face | Robotics, hardware and chips | In Edition 25, Saturday 14 Mar 2026

SOURCE Hugging Face, LeRobot v0.5.0: Scaling Every Dimension

Modular Diffusers

Modular Diffusers introduces a new way to build diffusion pipelines by composing reusable blocks, with a quickstart example running FLUX.2 Klein 4B and community pipelines including Krea Realtime Video achieving 11fps on a single B200 GPU.

Hugging Face | Models and APIs | In Edition 24, Saturday 7 Mar 2026

SOURCE Hugging Face, Introducing Modular Diffusers - Composable Building Blocks for Diffusion Pipelines

GGML and llama.cpp join HF

GGML, creators of llama.cpp, are joining Hugging Face, with Georgi Gerganov and team dedicating 100% of their time maintaining llama.cpp with full autonomy and leadership on technical directions and the community.

Hugging Face | Models and APIs | In Edition 22, Saturday 21 Feb 2026

SOURCE Hugging Face, GGML and llama.cpp join HF to ensure the long-term progress of Local AI

Gradio gr.HTML one-shot apps

Gradio 6 shipped gr.HTML with custom templates, scoped CSS, and JavaScript interactivity, letting Claude or any other frontier LLM generate frontend, backend, and state management in a single Python file with no build step.

Hugging Face | Tools and products | In Edition 22, Saturday 21 Feb 2026

SOURCE Hugging Face, One-Shot Any Web App with Gradio's gr.HTML

Unsloth and HF Jobs training

Unsloth and Hugging Face Jobs enable fast LLM fine-tuning of LiquidAI/LFM2.5-1.2B-Instruct through coding agents like Claude Code and Codex, with Unsloth providing approximately 2x faster training and about 60% less VRAM usage compared to standard methods.

Hugging Face | Models and APIs | In Edition 22, Saturday 21 Feb 2026

SOURCE Hugging Face, Train AI models with Unsloth and Hugging Face Jobs for FREE

CUDA kernel agent skill

Hugging Face released an agent skill that teaches Codex and Claude to write production CUDA kernels, producing an RMSNorm kernel for Qwen3-8B with an average 1.94x speedup on H100.

Hugging Face | Models and APIs | In Edition 21, Saturday 14 Feb 2026

SOURCE Hugging Face, Custom Kernels for All from Codex and Claude

OpenEnv Calendar Gym

Turing contributed a production-grade calendar management environment to OpenEnv, where agents achieved close to 90% success with explicit calendar identifiers but dropped to roughly 40% with natural language descriptions.

Hugging Face | Models and APIs | In Edition 21, Saturday 14 Feb 2026

SOURCE Hugging Face, OpenEnv in Practice: Evaluating Tool-Using Agents in Real-World Environments

Transformers.js v4

Transformers.js v4 is now available on NPM with a C++ WebGPU runtime, a 10x build time drop from 2 seconds to 200 milliseconds, and a default export that is 53% smaller.

Hugging Face | Models and APIs | In Edition 21, Saturday 14 Feb 2026

SOURCE Hugging Face, Transformers.js v4: Now Available on NPM!

Holo2-235B-A22B Preview

Holo2-235B-A22B Preview achieves 78.5% on Screenspot-Pro in agent mode within 3 steps and 79.0% on OSWorld G, and is available on Hugging Face.

Hugging Face | Models and APIs | In Edition 20, Saturday 7 Feb 2026

SOURCE Hugging Face, H Company's new Holo2 model takes the lead in UI Localization

SyGra Studio

SyGra 2.0.0 introduces Studio, an interactive environment that turns synthetic data generation into a visual canvas with guided model forms, data previews, and live execution streaming.

Hugging Face | Tools and products | In Edition 20, Saturday 7 Feb 2026

SOURCE Hugging Face, Introducing SyGra Studio

Daggr

Hugging Face released Daggr, an open-source Python library for building AI workflows that connect Gradio apps, ML models, and custom functions, with automatic visual canvas.

Hugging Face | Models and APIs | Published Thursday 29 Jan 2026

SOURCE Hugging Face, Introducing Daggr: Chain apps programmatically, inspect visually

AssetOpsBench

IBM Research released AssetOpsBench, a benchmark for industrial AI agents with 2.3M sensor telemetry points and 4.2K work orders.

Hugging Face | Models and APIs | Published Wednesday 21 Jan 2026

SOURCE Hugging Face, AssetOpsBench: Bridging the Gap Between AI Agent Benchmarks and Industrial Reality

Diff Transformer V2

Microsoft introduced Differential Transformer V2, improving inference speed and training stability for production LLMs, with experiments still running.

Hugging Face | Models and APIs | Published Tuesday 20 Jan 2026

SOURCE Hugging Face, Differential Transformer V2

Waypoint-1

Overworld released Waypoint-1, a real-time interactive video diffusion model trained on 10,000 hours of game footage, with weights on the Hub.

Hugging Face | Models and APIs | Published Tuesday 20 Jan 2026

SOURCE Hugging Face, Waypoint-1: Real-time Interactive Video Diffusion from Overworld

DGX Spark and Reachy Mini

NVIDIA showed how to build a personal office robot agent using DGX Spark with Reachy Mini, combining NVIDIA Nemotron 3 Nano for reasoning, Nemotron Nano 2 VL for vision, and ElevenLabs for text-to-speech.

Hugging Face | Robotics, hardware and chips | In Edition 16, Saturday 10 Jan 2026

SOURCE Hugging Face, NVIDIA brings agents to life with DGX Spark and Reachy Mini

Falcon-H1-Arabic

Falcon-H1-Arabic 3B, 7B, and 34B models outperform all SOTA models of similar sizes and sometimes bigger, with the 34B reaching approximately 75% on OALL and surpassing Llama-3.3-70B.

Hugging Face | Models and APIs | In Edition 16, Saturday 10 Jan 2026

SOURCE Hugging Face, Introducing Falcon-H1-Arabic: Pushing the Boundaries of Arabic Language AI with Hybrid Architecture

NVIDIA Cosmos Reason 2

NVIDIA released Cosmos Reason 2, an open reasoning vision language model for physical AI that tops the Physical AI Bench and Physical Reasoning leaderboards as the #1 open model for visual understanding.

Hugging Face | Models and APIs | In Edition 16, Saturday 10 Jan 2026

SOURCE Hugging Face, NVIDIA Cosmos Reason 2 Brings Advanced Reasoning To Physical AI

AprielGuard

ServiceNow released AprielGuard, an 8B parameter safety and security safeguard model that detects 16 safety risk categories and adversarial attacks including prompt injection, jailbreaks, chain-of-thought corruption, context hijacking, memory poisoning, and multi-agent exploit sequences across standalone prompts, multi-turn conversations, and agentic workflows.

Hugging Face | Models and APIs | In Edition 15, Saturday 27 Dec 2025

SOURCE Hugging Face, AprielGuard: A Guardrail for Safety and Adversarial Robustness in Modern LLM Systems

Transformers v5 tokenizers

Transformers v5 redesigns tokenizers, separating architecture from trained vocab, making them inspectable, customizable, and trainable from scratch.

Hugging Face | Models and APIs | Published Thursday 18 Dec 2025

SOURCE Hugging Face, Tokenization in Transformers v5: Simpler, Clearer, and More Modular

Nemotron 3 Nano evaluation

NVIDIA released the full evaluation recipe for Nemotron 3 Nano 30B A3B using NeMo Evaluator, with scores like 78.3 on MMLU-Pro and 89.1 on AIME 2025.

Hugging Face | Models and APIs | Published Wednesday 17 Dec 2025

SOURCE Hugging Face, The Open Evaluation Standard: Benchmarking NVIDIA Nemotron 3 Nano with NeMo Evaluator

CUGA on Hugging Face

CUGA, a configurable generalist agent, is now on Hugging Face Spaces, achieving #1 on AppWorld and top-tier on WebArena.

Hugging Face | Models and APIs | Published Monday 15 Dec 2025

SOURCE Hugging Face, CUGA on Hugging Face: Democratizing Configurable AI Agents

Codex HF skills

Codex can now run end-to-end ML experiments using Hugging Face Skills, including fine-tuning, evaluation, and report generation.

Hugging Face | Models and APIs | Published Thursday 11 Dec 2025

SOURCE Hugging Face, Codex is Open Sourcing AI models

llama.cpp router mode

llama.cpp server now supports router mode, allowing dynamic loading, unloading, and switching between multiple models without restarting.

Hugging Face | Models and APIs | Published Thursday 11 Dec 2025

SOURCE Hugging Face, New in llama.cpp: Model Management

DeepMath

Intel's DeepMath agent, built on Qwen3-4B Thinking and fine-tuned with GRPO, reduces output lengths by up to 66% while improving accuracy on MATH500, AIME, HMMT, and HLE benchmarks.

Hugging Face | Models and APIs | In Edition 12, Saturday 6 Dec 2025

SOURCE Hugging Face, DeepMath: A lightweight math reasoning Agent with smolagents

Transformers v5

Transformers v5.0.0rc-0 launches with over 400 model architectures, more than 750,000 compatible checkpoints on the Hub, and 1.2 billion total installs, up from 20,000 installs per day at v4.

Hugging Face | Models and APIs | In Edition 12, Saturday 6 Dec 2025

SOURCE Hugging Face, Transformers v5: Simple model definitions powering the AI ecosystem

swift-huggingface

swift-huggingface is a new Swift package providing a complete client for the Hugging Face Hub with Python-compatible cache.

Hugging Face | Models and APIs | Published Friday 5 Dec 2025

SOURCE Hugging Face, Introducing swift-huggingface: The Complete Swift Client for Hugging Face

FLUX.2

FLUX.2 is a new open image generation model from Black Forest Labs with a 32B parameter DiT and a single Mistral Small 3.1 text encoder. It supports multiple reference images and runs on 24GB GPUs with 4-bit quantization.

Hugging Face | In Edition 11, Saturday 29 Nov 2025

SOURCE Hugging Face, FLUX.2 - BFL's new open image generation model

OVHcloud Inference Provider

OVHcloud is now a supported Inference Provider on the Hugging Face Hub, offering serverless access to open-weight models like gpt-oss, Qwen3, DeepSeek R1, and Llama with pay-per-token pricing starting at €0.04 per million tokens.

Hugging Face | Models and APIs | In Edition 11, Saturday 29 Nov 2025

SOURCE Hugging Face, OVHcloud on Hugging Face Inference Providers

RapidFire AI

Hugging Face TRL now officially integrates with RapidFire AI to accelerate fine-tuning and post-training experiments, with internal benchmarks showing approximately 16-24x higher experimentation throughput than sequential config comparison.

Hugging Face | Models and APIs | In Edition 10, Saturday 22 Nov 2025

SOURCE Hugging Face, 20x Faster TRL Fine-tuning with RapidFire AI

AnyLanguageModel

AnyLanguageModel is a Swift package that provides a drop-in replacement for Apple's Foundation Models framework, supporting local and remote LLM providers including MLX, llama.cpp, and cloud APIs.

Hugging Face | Models and APIs | Published Thursday 20 Nov 2025

SOURCE Hugging Face, Introducing AnyLanguageModel: One API for Local and Remote LLMs on Apple Platforms

ROCm kernel builder

Hugging Face's kernels library now supports building and sharing ROCm kernels, with a guide using the RadeonFlow GEMM kernel for MI300X.

Hugging Face | Models and APIs | Published Monday 17 Nov 2025

SOURCE Hugging Face, Easily Build and Share ROCm Kernels with Hugging Face

Granite 4.0 Nano

IBM released Granite 4.0 Nano models from 350M to 1.5B parameters under Apache 2.0, trained on over 15T tokens with native support on vLLM, llama.cpp, and MLX.

Hugging Face | Models and APIs | In Edition 7, Saturday 1 Nov 2025

SOURCE Hugging Face, Granite 4.0 Nano: Just how small can you go?

Streaming datasets

Hugging Face improved streaming datasets with 100x fewer startup requests, 10x faster data resolution, and up to 2x faster streaming speed, outrunning local SSDs when training on 64xH100 with 256 workers.

Hugging Face | Models and APIs | In Edition 7, Saturday 1 Nov 2025

SOURCE Hugging Face, Streaming datasets: 100x More Efficient

huggingface_hub v1.0

huggingface_hub reached v1.0 with 113.5 million monthly downloads, powering access to over 2 million public models, 0.5 million public datasets, and 1 million public Spaces.

Hugging Face | Models and APIs | In Edition 7, Saturday 1 Nov 2025

SOURCE Hugging Face, huggingface_hub v1.0: Five Years of Building the Foundation of Open Machine Learning

LeRobot v0.4.0

LeRobot v0.4.0 introduces Datasets v3.0 with chunked episodes and streaming, integrates PI0, PI0.5, and GR00T N1.5 policies, adds LIBERO and Meta-World simulation support, and launches a plugin system for third-party hardware.

Hugging Face | Models and APIs | In Edition 6, Saturday 25 Oct 2025

SOURCE Hugging Face, LeRobot v0.4.0: Supercharging OSS Robot Learning

Sentence Transformers

Sentence Transformers is transitioning from the UKP Lab at TU Darmstadt to Hugging Face, with over 16,000 models on the Hub serving more than a million monthly unique users, and Tom Aarsen continuing as maintainer.

Hugging Face | Models and APIs | In Edition 6, Saturday 25 Oct 2025

SOURCE Hugging Face, Sentence Transformers is joining Hugging Face

VirusTotal scanning

Every one of the 2.2M+ public model and dataset repositories on the Hugging Face Hub is being continuously scanned with VirusTotal, comparing file hashes against its threat-intelligence database without sharing raw file contents.

Hugging Face | Models and APIs | In Edition 6, Saturday 25 Oct 2025

SOURCE Hugging Face, Hugging Face and VirusTotal collaborate to strengthen AI security

AI for Food Allergies

Hugging Face released Awesome Food Allergy Datasets, the first open collection of datasets on food allergies, to accelerate AI research.

Hugging Face | Models and APIs | Published Thursday 16 Oct 2025

SOURCE Hugging Face, AI for Food Allergies

GPT OSS on Intel Xeon

Intel and Hugging Face benchmarked GPT OSS on Google C4 VMs with Intel Xeon 6, finding a 1.7x TCO improvement over C3 VMs.

Hugging Face | Models and APIs | Published Thursday 16 Oct 2025

SOURCE Hugging Face, Google Cloud C4 Brings a 70% TCO improvement on GPT OSS with Intel and Hugging Face

Qwen3-8B agent speedup

OpenVINO.GenAI accelerates Qwen3-8B generation by about 1.4x on Intel Core Ultra using speculative decoding with a depth-pruned Qwen3-0.6B draft model.

Hugging Face | Models and APIs | In Edition 3, Saturday 4 Oct 2025

SOURCE Hugging Face, Accelerating Qwen3-8B Agent on Intel Core Ultra with Depth-Pruned Draft Models

RTEB benchmark

RTEB is a new retrieval embedding benchmark in beta that combines open and private datasets across 20 languages and domains like law, healthcare, code, and finance, using NDCG@10 as the default metric.

Hugging Face | Models and APIs | In Edition 3, Saturday 4 Oct 2025

SOURCE Hugging Face, Introducing RTEB: A New Standard for Retrieval Evaluation

dots.ocr on-device

Hugging Face converted dots.ocr, a 3B parameter model that surpasses Gemini 2.5 Pro on OmniDocBench, to Core ML and MLX, but the initial conversion is over 5GB and slow.

Hugging Face | Models and APIs | Published Thursday 2 Oct 2025

SOURCE Hugging Face, SOTA OCR with Core ML and dots.ocr

LeRobotDataset v3.0

LeRobotDataset v3.0 packs multiple episodes in a single file using relational metadata, and natively supports streaming mode to process large datasets on the fly without downloading prohibitively large collections onto disk.

Hugging Face | Models and APIs | In Edition 1, Saturday 20 Sep 2025

SOURCE Hugging Face, LeRobotDataset:v3.0: Bringing large-scale datasets to lerobot

Scaleway Inference Provider

Scaleway is now a supported Inference Provider on the Hugging Face Hub, offering serverless access to models like gpt-oss, Qwen3, DeepSeek R1, and Gemma 3 with pay-per-token pricing starting at €0.20 per million tokens.

Hugging Face | Models and APIs | In Edition 1, Saturday 20 Sep 2025

SOURCE Hugging Face, Scaleway on Hugging Face Inference Providers

Public AI Inference Provider

Public AI is now a supported Inference Provider on Hugging Face, offering free access to public and sovereign models like Apertus-70B, with usage free of charge at the time of writing.

Hugging Face | Models and APIs | Published Wednesday 17 Sep 2025

SOURCE Hugging Face, Public AI on Hugging Face Inference Providers

The edition every row came from arrives in your inbox, free.

Join free