How I automate my own job at Hugging Face using agents — Niels Rogge, Hugging Face
Thousands of GitHub issues, opened automatically, have produced exactly two negative replies. Niels Rogge works on what he calls the Google Drive to the hub team at Hugging Face, whose job is noticing that a paper's weights are sitting on Dropbox or Zenodo where nobody will find them, then asking the authors to publish on the hub instead. Hundreds of papers land on arXiv every day, so he automated himself.The useful part is that he built it twice, in opposite shapes, and explains why each time. The outreach half is a deterministic workflow: a model call at each step of the path he used to walk by hand, no agent framework at all, running nightly as a cron job on free GitHub Actions minutes, with tracing so he can inspect prompts, cost, and latency. He chose that because the prevailing advice when he built it was to avoid agents unless you genuinely need one. The follow up half, built recently, is the reverse. It is a fully autonomous loop whose main tool is bash, carrying one CLI, one skill, and a sandbox, fanned out so that every issue gets its own container. He is also candid that recipients are not told an agent wrote to them, on the grounds that it sends what he used to send himself and a disclosed bot tends to get closed unread.
Speaker info:
- https://x.com/NielsRogge
- https://www.linkedin.com/in/niels-rogge-a3b7a3127/
- https://nielsrogge.github.io/
Timestamps:
0:00 - The Google Drive to the hub problem
1:59 - Paper pages, metadata, and discoverability
3:41 - Why manual outreach does not scale
4:29 - The workflow he was running by hand
5:19 - Workflow or agent, and why it is not binary
7:03 - Nightly cron jobs, and tracing cost and latency
8:46 - The flood of replies, and automating follow up
9:36 - Switching to a fully autonomous loop
10:25 - Bash, one CLI, one skill, one sandbox
12:06 - A container per issue, fanned out
13:49 - What researchers actually reply
15:32 - Migrated models, and a 400 gigabyte dataset
18:06 - Open models, agents over workflows, and evaluation Receive SMS online on sms24.me
TubeReader video aggregator is a website that collects and organizes online videos from the YouTube source. Video aggregation is done for different purposes, and TubeReader take different approaches to achieve their purpose.
Our try to collect videos of high quality or interest for visitors to view; the collection may be made by editors or may be based on community votes.
Another method is to base the collection on those videos most viewed, either at the aggregator site or at various popular video hosting sites.
TubeReader site exists to allow users to collect their own sets of videos, for personal use as well as for browsing and viewing by others; TubeReader can develop online communities around video sharing.
Our site allow users to create a personalized video playlist, for personal use as well as for browsing and viewing by others.
@YouTubeReaderBot allows you to subscribe to Youtube channels.
By using @YouTubeReaderBot Bot you agree with YouTube Terms of Service.
Use the @YouTubeReaderBot telegram bot to be the first to be notified when new videos are released on your favorite channels.
Look for new videos or channels and share them with your friends.
You can start using our bot from this video, subscribe now to How I automate my own job at Hugging Face using agents — Niels Rogge, Hugging Face
What is YouTube?
YouTube is a free video sharing website that makes it easy to watch online videos. You can even create and upload your own videos to share with others. Originally created in 2005, YouTube is now one of the most popular sites on the Web, with visitors watching around 6 billion hours of video every month.