Hugging Face Agents Course

Team

university

https://bit.ly/hf-learn-agents

Activity Feed

AI & ML interests

None defined yet.

Recent Activity

burtenshaw updated a dataset 3 minutes ago

agents-course/certificates

Jofthomas updated a dataset about 1 hour ago

agents-course/unit4-students-scores

sergiopaniego updated a dataset about 10 hours ago

agents-course/final-certificates

View all activity

burtenshaw

updated a dataset 3 minutes ago

agents-course/certificates

Viewer • Updated 3 minutes ago • 42 • 14.4k • 64

Jofthomas

updated a dataset about 1 hour ago

agents-course/unit4-students-scores

Viewer • Updated about 1 hour ago • 6.82k • 12.1k • 9

sergiopaniego

updated 2 datasets about 10 hours ago

agents-course/final-certificates

Viewer • Updated about 10 hours ago • 5 • 1.08k • 4

agents-course/course-certificates-of-excellence

Viewer • Updated about 10 hours ago • 3.87k • 551 • 5

sergiopaniego

posted an update about 23 hours ago

Post

378

Want to get started with fine-tuning but don’t know where to begin? 🤓☝️

We’re expanding our collection of beginner-friendly free Colab notebooks so you can learn and fine-tune models using TRL at no cost

🔬 Check out the full list of free notebooks: https://huggingface.co/docs/trl/main/en/example_overview#notebooks

🔬 If you want more advanced content, we also have a lot to cover in the community tutorials: https://huggingface.co/docs/trl/community_tutorials

And now the obvious question: what would you like us to add next?

sergiopaniego

posted an update 3 days ago

Post

2173

NEW: @mistralai released a fantastic family of multimodal models, Ministral 3.

You can fine-tune them for free on Colab using TRL ⚡️, supporting both SFT and GRPO

Link to the notebooks:
- SFT: https://colab.research.google.com/github/huggingface/trl/blob/main/examples/notebooks/sft_ministral3_vl.ipynb
- GRPO: https://colab.research.google.com/github/huggingface/trl/blob/main/examples/notebooks/grpo_ministral3_vl.ipynb
- TRL and more examples: https://huggingface.co/docs/trl/index

2 replies

Jofthomas

posted an update 4 days ago

Post

3236

The new Mistral 3 models are here !

Today, we announce Mistral 3, the next generation of Mistral models. Mistral 3 includes three state-of-the-art small, dense models (14B, 8B, and 3B) and Mistral Large 3 – our most capable model to date – a sparse mixture-of-experts trained with 41B active and 675B total parameters.

All models are released under the Apache 2.0 license.

Ministrals :
https://huggingface.co/collections/mistralai/ministral-3

Mistral Large 3:
https://huggingface.co/collections/mistralai/mistral-large-3

2 replies

sergiopaniego

posted an update 4 days ago

Post

2078

ICYMI, transformers v5 is out!

Grab a coffee ☕ and go read the announcement blog https://huggingface.co/blog/transformers-v5

sergiopaniego

posted an update 5 days ago

Post

3039

want to use open models easily through an API?

Inference Providers might be exactly what you’re looking for sooo here’s a complete beginner-friendly walkthrough 🧐

https://www.youtube.com/watch?v=oxwsizy1Spw

2 replies

sergiopaniego

posted an update 9 days ago

Post

1682

nanochat is now in transformers!

The LLM by @karpathy is officially in the library, and we wrote a blog covering: how did we port the model, differences from the original, and how to run or train it.

go read it 🤓

nanochat-students/transformers

sergiopaniego

posted an update 11 days ago

Post

3922

you gotta go fast and go read the latest blog by @ror et al. explaining Continuous Batching in depth

https://huggingface.co/blog/continuous_batching

sergiopaniego

posted an update 12 days ago

Post

1693

Interested in RL training environments?

We just released a beginner-friendly walkthrough notebook!

Train a model to play Wordle using TRL + OpenEnv (TextArena) + GRPO + vLLM.

happy learning! 🌱

Notebook: https://github.com/huggingface/trl/blob/main/examples/notebooks/openenv_wordle_grpo.ipynb

OpenEnv guide in TRL: https://huggingface.co/docs/trl/main/en/openenv

sergiopaniego

posted an update 16 days ago

Post

315

Ya está disponible el vídeo de la charla del otro día en @nerdearla sobre IA abierta, por si queréis verla! 🤠

https://www.youtube.com/watch?v=p-JLn4xAkMw

1 reply

sergiopaniego

posted an update 17 days ago

Post

2568

we've just added several example scripts to TRL showing how to train models with GRPO using some of the new OpenEnv environments

train a model to interact with a browser (🎮 BrowserGym Env), play Wordle (🎮 Wordle Env) and moooore!

TRL (GRPO + vLLM) + OpenEnv! ⚡️

📝 go play with them: https://github.com/huggingface/trl/tree/main/examples/scripts/openenv

📝 examples list: https://huggingface.co/docs/trl/main/en/example_overview#scripts

sergiopaniego

posted an update 19 days ago

Post

1742

Who wants a TRL sticker? 🙋

https://github.com/huggingface/trl

1 reply

dawood

authored a paper 24 days ago

Gradio: Hassle-Free Sharing and Testing of ML Models in the Wild

Paper • 1906.02569 • Published Jun 6, 2019 • 1

sergiopaniego

posted an update about 1 month ago

Post

5360

fine-tuning a 14B model with TRL + SFT on a free Colab (T4 GPU)?
thanks to the latest TRL optimizations, you actually can!
sharing a new notebook showing how to do it 😎

colab: https://colab.research.google.com/github/huggingface/trl/blob/main/examples/notebooks/sft_trl_lora_qlora.ipynb

notebooks in TRL: https://github.com/huggingface/trl/tree/main/examples/notebooks

2 replies

sergiopaniego

posted an update about 1 month ago

Post

444

Gave a smol 🤏 intro to Agents using smolagents last Monday!
Sharing the slides in case you're curious. They serve as a gentle first step into the Agents Course we developed at @huggingface 🫶🫶

Course: https://huggingface.co/learn/agents-course/unit0/introduction

Workshop material: https://github.com/sergiopaniego/talks/tree/main/intro_to_agents

sergiopaniego

posted an update about 1 month ago

Post

3129

Sharing the slides from yesterday's talk about "Fine Tuning with TRL" from the @TogetherAgent x @huggingface workshop we hosted in our Paris office 🎃!

Link: https://github.com/sergiopaniego/talks/blob/main/fine_tuning_with_trl/Fine%20tuning%20with%20TRL%20(Oct%2025).pdf

sergiopaniego

posted an update about 1 month ago

Post

423

On-Policy distillation is trendy! and super useful!

HuggingFaceH4/on-policy-distillation

AI & ML interests

Recent Activity

Team members 9

agents-course's activity