← FeedLeaderboard →

Hugging Face

@huggingface

Last 8 weeks · 37 posts · 10 outlets writing · no. 4 of 20 by coverage

Window4w8w12w26w

Share of voice, week by week

29 Jun–5 Jul→ 17–23 Aug…
29 Jun–5 Jul6–12 Jul13–19 Jul20–26 Jul27 Jul–2 Aug3–9 Aug10–16 Aug17–23 Aug…
Postspeak 9/wk
Coveragepeak 10/wk

Who covers them

1 story of Hugging Face drew coverage in this window; a full bar covered it. A story counts once per outlet, however many pieces they filed.

  • Ars Technica1
  • Beyond Search1
  • Blockonomi1
  • Digital Trends1
  • Fortune1
  • Hello China Tech

Recent stories

Hugging FaceDeveloper tool17h ago

Same Cluster, 33 Points More Utilization: What Changed Was the Order

Hugging FaceDeveloper tool· 17h ago

Dharma-AI introduces a constraint-aware GPU allocator that improves GPU utilization by up to 33 percentage points and priority-weighted output by up to 105% compared to a FIFO scheduler, by…

Hugging FaceModel update4 days ago

State of Open Models: Summer 2026 Observations

1
  • iThinkDifferent1
  • Memeburn1
  • PYMNTS1
  • Silicon Republic1
  • Tech in Asia1
  • The Standard1
  • Hugging FaceModel update· 4 days ago

    In the first seven months of 2026, the largest open models from Chinese labs ranged from 754B to 2.78 trillion parameters, while U.S.

    Hugging FaceProduct launch4 days ago

    Record, train, and deploy from one place with Strands Agents, LeRobot, and Hugging Face Storage Buckets

    Hugging FaceProduct launch· 4 days ago

    Hugging Face and AWS announce a streaming data loop for Strands Robots that records, trains, and deploys robot policies using Hugging Face Storage Buckets and LeRobot format, enabling continuous…

    Hugging FaceProduct launch5 days ago

    What We Learned by Reproducing 2,200 papers from ICML

    Hugging FaceProduct launch· 5 days ago

    Hugging Face ran a hackathon where 1,221 community members used coding agents to reproduce 2,226 papers from ICML 2026, finding that 51% had at least one verified claim and 23% had at least one…

    Hugging FaceProduct launch5 days ago

    Introducing OlmoEarth embeddings: Custom embedding exports from OlmoEarth Studio for downstream analysis

    Hugging FaceProduct launch· 5 days ago

    OlmoEarth Studio now supports custom embedding exports, allowing users to compute and download embedding vectors from OlmoEarth foundation models for downstream tasks like similarity search and…

    Hugging Face5 days ago

    LFM2.5-VL-3B for Better and Faster Vision Capabilities for the Edge

    Hugging Face· 5 days ago

    LFM2.5-VL-3B is our most capable vision-language model you can run on your own hardware. It understands documents and screens alike, grounds objects, and can call tools.

    Hugging Face6 days ago

    Thinking of ACE? We Can Do It with Fewer Tokens

    Hugging Face· 6 days ago

    ALTK-Evolve and ACE both let an agent learn from its own trajectories. The difference is what they do with what they learn — and that decides the token bill.

    Hugging Face7 days ago

    Build Low-Latency Multilingual Voice Agents: Open Weights & Full Deployment Control with NVIDIA Magpie TTS

    Hugging Face· 7 days ago

    Every voice interaction has a latency budget. By the time a user hears your application respond, you've already spent precious milliseconds capturing audio, transcribing speech, running an LLM,…

    Hugging Face8 days ago

    Making Knowledge Distillation Cheap Enough to Run at Scale

    Hugging Face· 8 days ago

    Knowledge distillation , training a smaller student model to match the performance of a larger teacher, is a well-known technique in Machine Learning.

    Hugging Face8 days ago

    Meta is back with Muse Glimmer: local, agentic, multimodal, and open source

    Hugging Face· 8 days ago

    ! Published August 10, 2026 Update on GitHub Upvote 88 +82 Pedro Cuenca pcuenq Follow merve merve Follow ben burtenshaw burtenshaw Follow Aritra Roy Gosthipaty ariG23498 Follow Great news from the…

    Hugging FaceNew model10 days ago

    TutorMoments: Do AI tutors know when to help and when to hold back?

    Hugging FaceNew model· 10 days ago

    Hugging Face and Allen AI introduce TutorMoments, a framework to evaluate whether LLMs can balance helping students versus letting them struggle, using real tutoring transcripts and teacher…

    Hugging FaceInfrastructure12 days ago

    Baseten on Hugging Face Inference Providers 🔥

    Hugging FaceInfrastructure· 12 days ago

    Baseten is now a supported Inference Provider on the Hugging Face Hub, offering serverless inference for conversational and text-generation tasks with models like Kimi K3, DeepSeek V4 Flash, and…

    Hugging Face18 days ago

    GPU Management: Why Idle GPUs Are the New Grounded Aircraft

    Hugging Face· 18 days ago

    Utilization, not intelligence, is the next real constraint in AI. Aviation learned this the hard way. For most of the industry's history, the number that best predicted whether an airline would…

    Hugging Face20 days ago

    The OlmoEarth Platform: Geospatial inference at planetary scale

    Hugging Face· 20 days ago

    🌍 Learn more about OlmoEarth Platform: https://allenai.org/olmoearth The OlmoEarth models are our family of Earth observation foundation models, pretrained on roughly 10 terabytes of multimodal…

    Hugging Face20 days ago

    LFM2.5-Encoders for Fast Long-Context Inference on CPU

    Hugging Face· 20 days ago

    Today, we release two new encoder models on Hugging Face: LFM2.5-Encoder-230M and LFM2.5-Encoder-350M . They match the quality of larger models but stay fast as inputs get longer.

    Hugging Face22 days ago

    NVIDIA Cosmos-H-Dreams: Bringing Real-Time Generative Simulation to Surgical Robotics

    Hugging Face· 22 days ago

    Surgical robotics is moving quickly from teleoperation toward increasingly capable vision-language-action policies. But evaluating and training these systems remains difficult.

    Hugging Face22 days ago

    Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident

    Hugging Face· 22 days ago

    A companion technical writeup to our incident disclosure . This post walks through how the intrusion actually worked: the two initial-access vectors, how the agent pivoted and moved laterally,…

    Hugging Face26 days ago

    Bringing Nunchaku 4-bit Diffusion Inference to Diffusers

    Hugging Face· 26 days ago

    Large diffusion transformers can create stunning images (or even videos, audio snippets, and now text), but loading a modern text-to-image model in BF16 precision often requires 20-30 GB of VRAM,…

    Hugging Face28 days ago

    Grabette: an open system to record robot-manipulation data

    Hugging Face· 28 days ago

    . And build a shared dataset, together. Published July 21, 2026 Update on GitHub Upvote 49 +43 Steve Nguyen SteveNguyen Follow pollen-robotics Claire Houziel chouziel Follow pollen-robotics Gaelle…

    Hugging Face16 Jul

    Newer Models, Same Advantage

    Hugging Face· 16 Jul

    Despite newer architectures, DharmaOCR outperformed Mistral OCR4 and Unlimited-OCR on Brazilian Portuguese through domain specialization and targeted training.

    Hugging Face16 Jul

    Security incident disclosure — July 2026

    Hugging Face· 16 Jul

    Earlier this week, we detected and responded to an intrusion into part of our production infrastructure.

    Hugging Face15 Jul

    Model Routing Is Simple. Until It Isn’t.

    Hugging Face· 15 Jul

    Building a router into your agent sounds like an easy win. Send simple requests to cheaper models, reserve expensive ones for harder tasks, or route by specialty — Claude for code, Gemini for…

    Hugging Face15 Jul

    Introducing Real World VoiceEQ: Measuring the human quality of voice AI

    Hugging Face· 15 Jul

    Existing benchmarks suggest voice AI is nearing human-level performance but real-world conversations tell a different story. Voice is rapidly becoming AI's primary interface.

    Hugging Face15 Jul

    Welcome Inkling by Thinking Machines

    Hugging Face· 15 Jul

    Inkling now comes in smaller size 🤗 Inkling-Small is out by Thinking Machines Lab. We have updated this post with performance , and deployment configurations for the Inkling-Small and the…

    Hugging Face10 Jul

    Profiling in PyTorch (Part 3): Attention is all you profile

    Hugging Face· 10 Jul

    This is the third post of Profiling in PyTorch, a series where we slowly build the skill of reading profiler traces and use it to drive optimization Profiling in PyTorch (Part 1): A Beginner's Guide…

    Hugging Face8 Jul

    Native-speed vLLM transformers modeling backend

    Hugging Face· 8 Jul

    TL;DR : The transformers vLLM backend is now as fast (or faster) than custom vLLM implementations for many LLM architectures.

    Hugging Face7 Jul

    From Hugging Face to Amazon SageMaker Studio in one click

    Hugging Face· 7 Jul

    Today, we’re excited to announce a deep-link integration between Hugging Face and Amazon SageMaker AI .

    Hugging Face7 Jul

    Hugging Face Models on Foundry Managed Compute

    Hugging Face· 7 Jul

    At Microsoft Build 2026, we announced Foundry Managed Compute and Hugging Face models on Foundry — a curated catalog of open-weight models from the Hugging Face ecosystem, refreshed weekly,…

    Hugging Face7 Jul

    LeRobot v0.6.0: Imagine, Evaluate, Improve

    Hugging Face· 7 Jul

    This new release is about closing the robot learning loop: policies that imagine the future before acting, reward models that tell you when your robot succeeds, a deployment CLI that turns failures…

    Hugging Face7 Jul

    Run AI workloads on any cloud, store on Hugging Face: zero-egress storage with SkyPilot

    Hugging Face· 7 Jul

    For most teams, models and datasets live in a bucket in one region of one cloud. The GPUs you can get, whether for development, training, or serving, increasingly sit on a different cloud than your…

    Hugging Face6 Jul

    PRX Part 4: Our Data Strategy

    Hugging Face· 6 Jul

    Welcome back! This is Part 4 of the PRX series. Parts 1 to 3 covered model architectures , training design , and a 24-hour speedrun .

    Hugging Face6 Jul

    🤗 Kernels: Major Updates

    Hugging Face· 6 Jul

    In our previous post (From Zero to GPU) , we introduced the 🤗 Kernels project, which aims at standardizing how custom kernels are packaged, distributed, and consumed.

    Hugging Face1 Jul

    Hugging Face and Cerebras bring Gemma 4 to real-time voice AI

    Hugging Face· 1 Jul

    For voice AI, latency is a critical parameter. Developers have made tremendous progress in model quality, but the user experience is still often limited by response times.

    Hugging Face30 Jun

    ScarfBench: Benchmarking AI Agents for Enterprise Java Framework Migration

    Hugging Face· 30 Jun

    Modernizing enterprise applications is one of the largest and most expensive software engineering activities organizations undertake.

    Hugging Face30 Jun

    Why Specialization Is Inevitable

    Hugging Face· 30 Jun

    What optimization theory, evolutionary biology, competitive markets, and machine learning all predict — and why the answer is the same --- Those who follow Dharma AI already know that we view…

    Hugging Face30 Jun

    Featuring Every Eval Ever Results on Hugging Face Model Pages

    Hugging Face· 30 Jun

    Every Eval Ever (EEE) and Hugging Face Community Evals are now intercompatible. We enable cross-posting and interpreting evaluation results, while linking to open models, leaderboards, and a unified…

    Hugging Face29 Jun

    DiScoFormer: One transformer for density and score, across distributions

    Hugging Face· 29 Jun

    Many problems in machine learning and the sciences come down to the same task: you have a collection of data points and want to recover the distribution they came from—which values are common, and…

    Hugging Face26 Jun

    Run a vLLM Server on HF Jobs in One Command

    Hugging Face· 26 Jun

    You can spin up a private, OpenAI-compatible LLM endpoint on Hugging Face infrastructure with a single command — no servers to provision, no Kubernetes, pay-per-second.

    Hugging Face24 Jun

    Accelerating Transformers Fine-Tuning with NVIDIA NeMo AutoModel

    Hugging Face· 24 Jun

    HuggingFace Transformers has become the foundation of the open-source AI ecosystem, and the recent Transformers v5 release strengthened it with first-class support for Mixture-of-Experts (MoE)…

    Hugging Face24 Jun

    Introducing the FFASR Leaderboard: Benchmarking ASR in the Real World

    Hugging Face· 24 Jun

    🚀 First open far-field ASR benchmark: community-driven evaluation across 14 simulated rooms, validated against real-world measurements: https://huggingface.co/spaces/treble-technologies/ffasr 📉 The…

    Hugging Face23 Jun

    Experimenting with the proposed Cross-Origin Storage API in Transformers.js

    Hugging Face· 23 Jun

    (This is a guest post by Developer Relations Engineer Thomas Steiner from the Chrome team at Google.) Transformers.js provides Web developers with a simple way to use the power of transformers in…

    Hugging Face23 Jun

    Shipping huggingface_hub every week with AI, open tools, and a human in the loop

    Hugging Face· 23 Jun

    huggingface_hub is the Python client at the base of the Hugging Face ecosystem. transformers , datasets , diffusers , sentence-transformers and dozens of other libraries depend on it to talk to the…

    Hugging Face22 Jun

    PP-OCRv6 on Hugging Face: 50-Language OCR from 1.5M to 34.5M Parameters

    Hugging Face· 22 Jun

    Evaluate PP-OCRv6 online, then integrate lightweight, production-ready OCR with PaddlePaddle, Transformers, or ONNX Runtime backend.

    Hugging Face22 Jun

    We got local models to triage the OpenClaw repo for FREE!*

    Hugging Face· 22 Jun

    *Free as in beer, excluding the cost of electricity, and assuming you already own the hardware June 2026 will go down as the moment that people realized closed models can be taken away.

    Hugging Face18 Jun

    MosaicLeaks: Can your research agent keep a secret?

    Hugging Face· 18 Jun

    Deep research agents increasingly combine private local documents with external tools like web retrieval, creating a privacy risk: an agent's external queries may leak sensitive information.

    Hugging Face18 Jun

    Is it agentic enough? Benchmarking open models on your own tooling

    Hugging Face· 18 Jun

    Benchmarking transformers revisions across different metrics This is a human-made, agent-focused blogpost.

    Hugging Face18 Jun

    Beyond LoRA: Can you beat the most popular fine-tuning technique?

    Hugging Face· 18 Jun

    When you plan to fine-tune a model in a parameter-efficient way, think beyond LoRA If you want to fine-tune an open model on your own data, you are probably interested in so-called…

    Hugging Face17 Jun

    From the Hugging Face Hub to robot hardware with Strands Agents and LeRobot

    Hugging Face· 17 Jun

    A walkthrough of the LeRobot integration in Strands Robots - one agent loop, from a Hub dataset to a physical robot, with sim-to-real datasets in the same on-disk format and policies you swap with a…

    Hugging Face17 Jun

    GLM-5.2: Built for Long-Horizon Tasks

    Hugging Face· 17 Jun

    We're introducing GLM-5.2, our latest flagship model for long-horizon tasks. It marks a substantial leap in long-horizon task capability over its predecessor GLM-5.1 and, for the first time, delivers…

    Hugging Face17 Jun

    Agentic Resource Discovery: Let agents search

    Hugging Face· 17 Jun

    for tools, skills, and other agents. Published June 17, 2026 Update on GitHub Upvote 21 +15 ben burtenshaw burtenshaw Follow shaun smith evalstate Follow If you build with agents today, you probably…