Schneider Marcell

Sneccello

AI & ML interests

None yet

Recent Activity

reacted to jasoncorkill's post with 🔥 2 days ago

🚀 Rapidata: Setting the Standard for Model Evaluation Rapidata is proud to announce our first independent appearance in academic research, featured in the Lumina-Image 2.0 paper. This marks the beginning of our journey to become the standard for testing text-to-image and generative models. Our expertise in large-scale human annotations allows researchers to refine their models with accurate, real-world feedback. As we continue to establish ourselves as a key player in model evaluation, we’re here to support researchers with high-quality annotations at scale. Reach out to info@rapidata.ai to see how we can help. https://huggingface.co/papers/2503.21758

liked a dataset 10 days ago

Rapidata/OpenAI-4o_t2i_human_preference

reacted to jasoncorkill's post with 🧠 19 days ago

At Rapidata, we compared DeepL with LLMs like DeepSeek-R1, Llama, and Mixtral for translation quality using feedback from over 51,000 native speakers. Despite the costs, the performance makes it a valuable investment, especially in critical applications where translation quality is paramount. Now we can say that Europe is more than imposing regulations. Our dataset, based on these comparisons, is now available on Hugging Face. This might be useful for anyone working on AI translation or language model evaluation. https://huggingface.co/datasets/Rapidata/Translation-deepseek-llama-mixtral-v-deepl

View all activity

Organizations

Sneccello's activity

reacted to jasoncorkill's post with 🔥 2 days ago

Post

2263

🚀 Rapidata: Setting the Standard for Model Evaluation

Rapidata is proud to announce our first independent appearance in academic research, featured in the Lumina-Image 2.0 paper. This marks the beginning of our journey to become the standard for testing text-to-image and generative models. Our expertise in large-scale human annotations allows researchers to refine their models with accurate, real-world feedback.

As we continue to establish ourselves as a key player in model evaluation, we’re here to support researchers with high-quality annotations at scale. Reach out to info@rapidata.ai to see how we can help.

Lumina-Image 2.0: A Unified and Efficient Image Generative Framework (2503.21758)

liked a dataset 10 days ago

Rapidata/OpenAI-4o_t2i_human_preference

Viewer • Updated 8 days ago • 13k • 1.18k • 29

reacted to jasoncorkill's post with 🧠 19 days ago

Post

3795

At Rapidata, we compared DeepL with LLMs like DeepSeek-R1, Llama, and Mixtral for translation quality using feedback from over 51,000 native speakers. Despite the costs, the performance makes it a valuable investment, especially in critical applications where translation quality is paramount. Now we can say that Europe is more than imposing regulations.

Our dataset, based on these comparisons, is now available on Hugging Face. This might be useful for anyone working on AI translation or language model evaluation.

Rapidata/Translation-deepseek-llama-mixtral-v-deepl

1 reply

reacted to jasoncorkill's post with 🔥 25 days ago

Post

2205

Benchmarking Google's Veo2: How Does It Compare?

The results did not meet expectations. Veo2 struggled with style consistency and temporal coherence, falling behind competitors like Runway, Pika, Tencent, and even Alibaba. While the model shows promise, its alignment and quality are not yet there.

Google recently launched Veo2, its latest text-to-video model, through select partners like fal.ai. As part of our ongoing evaluation of state-of-the-art generative video models, we rigorously benchmarked Veo2 against industry leaders.

We generated a large set of Veo2 videos spending hundreds of dollars in the process and systematically evaluated them using our Python-based API for human and automated labeling.

Check out the ranking here: https://www.rapidata.ai/leaderboard/video-models

Rapidata/text-2-video-human-preferences-veo2

liked 2 datasets 26 days ago

Rapidata/text-2-video-human-preferences-veo2

Viewer • Updated 26 days ago • 760 • 488 • 12

Rapidata/text-2-video-human-preferences-wan2.1

Viewer • Updated 26 days ago • 787 • 579 • 16

liked a dataset 30 days ago

Rapidata/Translation-deepseek-llama-mixtral-v-deepl

Viewer • Updated 27 days ago • 845 • 405 • 15

updated a dataset 30 days ago

Rapidata/text-2-image-Rich-Human-Feedback

Viewer • Updated 30 days ago • 13k • 553 • 36

reacted to jasoncorkill's post with 🚀 about 2 months ago

Post

2553

This dataset was collected in roughly 4 hours using the Rapidata Python API, showcasing how quickly large-scale annotations can be performed with the right tooling!

All that at less than the cost of a single hour of a typical ML engineer in Zurich!

The new dataset of ~22,000 human annotations evaluating AI-generated videos based on different dimensions, such as Prompt-Video Alignment, Word for Word Prompt Alignment, Style, Speed of Time flow and Quality of Physics.

Rapidata/text-2-video-Rich-Human-Feedback

1 reply

reacted to jasoncorkill's post with ❤️ about 2 months ago

Post

4625

Runway Gen-3 Alpha: The Style and Coherence Champion

Runway's latest video generation model, Gen-3 Alpha, is something special. It ranks #3 overall on our text-to-video human preference benchmark, but in terms of style and coherence, it outperforms even OpenAI Sora.

However, it struggles with alignment, making it less predictable for controlled outputs.

We've released a new dataset with human evaluations of Runway Gen-3 Alpha: Rapidata's text-2-video human preferences dataset. If you're working on video generation and want to see how your model compares to the biggest players, we can benchmark it for you.

🚀 DM us if you’re interested!

Dataset: Rapidata/text-2-video-human-preferences-runway-alpha

1 reply

liked a dataset about 2 months ago

Rapidata/text-2-video-human-preferences-luma-ray2

Viewer • Updated Feb 11 • 948 • 78 • 10

liked a dataset 2 months ago

Rapidata/text-2-video-Rich-Human-Feedback

Viewer • Updated Feb 7 • 198 • 212 • 13

upvoted a collection 2 months ago

Sora Rich annotation

Collection

Contains a rich set of different types of annotations for a set of 198 videos • 6 items • Updated Feb 4 • 10

liked 4 datasets 2 months ago

reacted to jasoncorkill's post with 🚀 2 months ago

Post

2724

We benchmarked @xai-org 's Aurora model, as far as we know the first public evaluation of the model at scale.

We collected 401k human annotations in over the past ~2 days for this, we have uploaded all of the annotation data here on huggingface with a fully permissive license
Rapidata/xAI_Aurora_t2i_human_preferences

1 reply

liked 2 datasets 3 months ago

Rapidata/image-preference-demo

Viewer • Updated Jan 10 • 200 • 114 • 12

Rapidata/flux1.1-likert-scale-preference

Viewer • Updated Jan 10 • 1.12k • 100 • 14