Olivier Dehaene's picture

Olivier Dehaene

olivierdehaene

AI & ML interests

None yet

Recent Activity

Organizations

BigScience Workshop's profile picture OpenAssistant's profile picture LLHF's profile picture SLLHF's profile picture blhf's profile picture

olivierdehaene's activity

New activity in Alibaba-NLP/gte-Qwen2-1.5B-instruct 7 months ago
New activity in sanchit-gandhi/whisper-jax 7 months ago

Fix Dockerfile

1
#127 opened 7 months ago by
olivierdehaene
New activity in mistralai/Mistral-Nemo-Instruct-2407 8 months ago

model is not working

1
#74 opened 8 months ago by
lowpex
replied to mayank-mishra's post about 1 year ago
view reply

Nice blog!
@osanseviero we have been doing this in TGI and TEI for a while ;)
Padding free implementations also make dynamic batching easier to implement and more predictable in memory.

reacted to loubnabnl's post with β€οΈπŸ€―πŸ€— about 1 year ago
view post
Post
⭐ Today we’re releasing The Stack v2 & StarCoder2: a series of 3B, 7B & 15B code generation models trained on 3.3 to 4.5 trillion tokens of code:

- StarCoder2-15B matches or outperforms CodeLlama 34B, and approaches DeepSeek-33B on multiple benchmarks.
- StarCoder2-3B outperforms StarCoderBase-15B and similar sized models.
- The Stack v2 a 4x larger dataset than the Stack v1, resulting in 900B unique code tokens πŸš€
As always, we released everything from models and datasets to curation code. Enjoy!

πŸ”— StarCoder2 collection: bigcode/starcoder2-65de6da6e87db3383572be1a
πŸ”— Paper: https://drive.google.com/file/d/17iGn3c-sYNiLyRSY-A85QOzgzGnGiVI3/view
πŸ”— BlogPost: https://huggingface.co/blog/starcoder2
πŸ”— Code Leaderboard: bigcode/bigcode-models-leaderboard
published an article over 1 year ago
view article
Article

Welcome Mixtral - a SOTA Mixture of Experts on Hugging Face

By lewtun and 6 others β€’
β€’ 12
New activity in BAAI/bge-reranker-large over 1 year ago

Add fast tokenizer

1
#4 opened over 1 year ago by
olivierdehaene
New activity in BAAI/bge-reranker-base over 1 year ago

Add fast tokenizer

1
#5 opened over 1 year ago by
olivierdehaene
New activity in HuggingFaceH4/zephyr-chat over 1 year ago
New activity in thenlper/gte-base over 1 year ago
New activity in llmrails/ember-v1 over 1 year ago
New activity in BAAI/bge-large-en-v1.5 over 1 year ago
New activity in BAAI/bge-base-en-v1.5 over 1 year ago
New activity in tiiuae/falcon-7b-instruct almost 2 years ago

Add hf endpoint handler.py

#24 opened almost 2 years ago by
olivierdehaene