Curious 2B

Curious 2B is the on-device model used by CuriousLM. It is Qwen3.5 2B fine-tuned on the tasks CuriousLM's assistant performs on Android: answering questions from the user's calendar, adding calendar events through a tool call, reading incoming messages and notifications (what a message asks, whether a reply is needed, bills, bookings, deliveries and scams), and writing the assistant's morning and evening notes and reply drafts.

The model runs locally with llama.cpp. CuriousLM supplies the calendar and notification data and runs the tools; nothing is sent to a server.

Releases

Release File Size SHA-256
curious-2b-261010 (10 October 2026, current) curious-2b-261010-Q4_0.gguf 1.21 GB e106382634f3ec5319fb7fa5a6d0bd28e8c5fb1d031d8ae29b8e8866fc743f0f
curious-2b-261009 (9 October 2026) curious-2b-261009-Q4_0.gguf 1.21 GB 42b4c223434842288c999a447472df33d118169b60eeda55622d0c01866ca333
curious-2b-2610 (8 October 2026) curious-2b-2610-Q4_0.gguf 1.21 GB 45840b88cc63279df6b42c857832068d85c6f9a13745cebcb64fc792af45320d

Evaluation

Held-out cases, both releases under the same prompts and settings.

Phone and calendar

261009 261010
Turns passed (665) 88.0% 90.1%
Requests with times said in words, marked and cleared days, day of the month, people in titles (580) 58.6% 79.8%
Questions about a past day (150) 56.7% 75.3%
Reminder and alarm times in English, Spanish, French and German, from MTOP (1,050) 83.7% 88.1%
Says it cannot read the phone 1.1% 0.8%
Claims an action that was not taken 0.5% 0.3%

Notification reading

261009 261010
Standard set (519 tasks) 96.5% 98.5%
Hard set: receipts, deliveries, friends' plans, scams 93.5% 95.2%
Real text messages (300) 53.7% 78.7%
Real scams flagged (of 100) 94 99
Promotions flagged as scams (of 100) 47 21

The assistant's writing

261009 261010
Morning and evening notes, what came in, reply drafts 54.0% 61.2%
Home briefing leads with what needs the user (120) 40.8% 58.3%

General chat

261009 261010
Knowledge questions 93.0% 94.5%
Identifies itself correctly 94.4% 98.6%
Judged false claims per open answer (105) 2.6 2.9

Use

In CuriousLM: Models, then Curious 2B.

With llama.cpp:

llama-server -m curious-2b-261010-Q4_0.gguf --jinja

The tool definitions and prompts the model was tuned on are CuriousLM's own. For general function calling outside the app, Qwen3.5 2B scores higher.

Details

Base model Qwen3.5 2B
Method LoRA fine-tune, merged
Format GGUF, Q4_0
Languages English, Spanish, French, Portuguese, German, Italian, Dutch, Chinese

Licence

Apache 2.0. Fine-tuned from Qwen3.5-2B, Copyright 2026 Alibaba Cloud, Apache 2.0. Training data includes MASSIVE (Amazon) and When2Call (NVIDIA), both CC BY 4.0; the SMS Phishing Dataset (Mishra and Soni, Mendeley Data, doi:10.17632/f45bkkt8pr.1) and the SMS Spam Collection (Almeida and Hidalgo, UCI), both CC BY 4.0; and spoken date and time phrases from Recognizers-Text (Microsoft, MIT) and Duckling (Meta, BSD 3-Clause).

Downloads last month
11
GGUF
Model size
2B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for ebenezerdon/curious-2b

Finetuned
Qwen/Qwen3.5-2B
Finetuned
(475)
this model