AI & ML interests

None defined yet.

Recent Activity

Articles

CohereForAI's activity

davanstrien 
posted an update about 22 hours ago
view post
Post
1446
Hacked together a way to log trl GRPO training completions to a 🤗 dataset repo. This allows you to:

- Track rewards from multiple reward functions
- Treat the completion and rewards from training as a "proper" dataset and do EDA
- Share results for open science

The implementation is super hacky, but I'm curious if people would find this useful.

To push completions to the Hub, you just need two extra parameters:

log_completions=True
log_completions_hub_repo='your-username/repo-name'

Example dataset: davanstrien/test-logs
Colab: https://colab.research.google.com/drive/1wzBFPVthRYYTp-mEYlznLg_e_0Za1M3g

Brittawnya 
updated a Space 3 days ago
davanstrien 
posted an update 5 days ago
davanstrien 
posted an update 7 days ago
view post
Post
1825
How do you make 1M+ Hugging Face models & datasets more discoverable?

davanstrien/Smol-Hub-tldr!

I fine-tuned HuggingFaceTB/SmolLM2-360M to generate one-line summaries from a model or dataset README.

Its own self-description?
"A model for generating concise summaries of model & dataset cards from the Hugging Face Hub"

The goal? Make it easier to find the right models and datasets for your specific needs. It's already powering a semantic search for datasets Space.

It's still a WIP but thanks to @loubnabnl , @anton-l , @eliebak et al, for cooking such a nice base model for fine-tuning small, efficient models for specific domains and tasks. 🙏
davanstrien 
posted an update 8 days ago
davanstrien 
posted an update 23 days ago
davanstrien 
posted an update 24 days ago
davanstrien 
posted an update 25 days ago
view post
Post
2019
🌍 Big step for multilingual AI data!

The Hugging Face community has rated educational content in languages spoken by 1.6 billion people! New additions:
• Japanese
• Italian
• Old High German

Learn more and contribute: https://huggingface.co/blog/davanstrien/fineweb2-community

These ratings can help enhance training data for major world languages.
  • 1 reply
·

Web search fails

1
#46 opened about 1 month ago by
kaleidoskop-hug

rrr

#42 opened 3 months ago by
kazu001

Adding Evaluation Results

1
#12 opened about 1 month ago by
T145

-snip-

#11 opened about 2 months ago by
LPN64