Instructions to use BoomJules/molly-quantum-software-architect with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- PEFT
How to use BoomJules/molly-quantum-software-architect with PEFT:
from peft import PeftModel from transformers import AutoModelForCausalLM base_model = AutoModelForCausalLM.from_pretrained("meta-llama/Llama-3.1-8B-Instruct") model = PeftModel.from_pretrained(base_model, "BoomJules/molly-quantum-software-architect") - Notebooks
- Google Colab
- Kaggle
Molly Specialist β Quantum Software Architect
Generates correct Qiskit and Cirq circuit code, identifies gate decomposition errors, and explains quantum algorithm trade-offs more accurately than the base model.
Part of Molly, an orchestrator that keeps a library of small domain specialists over one quantized base and routes each request to the right one, so a single machine answers across many fields without loading a separate large model for each.
What this specialist handles well
- Translates quantum algorithms into optimized Qiskit or Cirq circuit code
- Identifies errors in quantum gate decompositions and circuit depth optimization
- Explains quantum error correction scheme trade-offs for specific hardware topologies
Try it with
- "How do I implement Shor's algorithm in Qiskit with minimal circuit depth?"
- "What error correction code works best for a 127-qubit heavy-hex topology?"
- "Convert this QAOA circuit from Cirq to Qiskit preserving gate fidelity"
Before you run: the base model is gated
This adapter needs the base weights, and the base is access-gated. Do this once:
- Accept the base licence: https://huggingface.co/meta-llama/Llama-3.1-8B-Instruct
- Create a read token: https://huggingface.co/settings/tokens
- Make the token available:
- Google Colab: Secrets panel (key icon) β Add new secret β name
HF_TOKEN, enable Notebook access. - Kaggle: Add-ons β Secrets β add
HF_TOKEN. - Local:
huggingface-cli loginorexport HF_TOKEN=...
- Google Colab: Secrets panel (key icon) β Add new secret β name
Skipping this gives GatedRepoError / 401 Unauthorized when the base loads. A stored
Colab secret is not applied automatically β authenticate in code, as below.
Quickstart
# pip install -U transformers peft accelerate
import os, torch
from huggingface_hub import login
try:
from google.colab import userdata
login(userdata.get("HF_TOKEN"))
except Exception:
tok = os.environ.get("HF_TOKEN")
login(tok) if tok else login()
from transformers import AutoModelForCausalLM, AutoTokenizer
from peft import PeftModel
BASE = "meta-llama/Llama-3.1-8B-Instruct"
ADAPTER = "BoomJules/molly-quantum-software-architect"
tok = AutoTokenizer.from_pretrained(BASE)
base = AutoModelForCausalLM.from_pretrained(BASE, torch_dtype=torch.bfloat16, device_map="auto")
model = PeftModel.from_pretrained(base, ADAPTER).eval()
msgs = [{"role": "user", "content": "Your question here"}]
ids = tok.apply_chat_template(msgs, add_generation_prompt=True, return_tensors="pt").to(model.device)
out = model.generate(ids, max_new_tokens=300)
print(tok.decode(out[0][ids.shape[1]:], skip_special_tokens=True))
Low-VRAM (4-bit) β fits a free Colab/Kaggle GPU (~6β7 GB)
# pip install -U transformers peft accelerate bitsandbytes
import os, torch
from huggingface_hub import login
try:
from google.colab import userdata
login(userdata.get("HF_TOKEN"))
except Exception:
login(os.environ.get("HF_TOKEN"))
from transformers import AutoModelForCausalLM, AutoTokenizer, BitsAndBytesConfig
from peft import PeftModel
bnb = BitsAndBytesConfig(load_in_4bit=True, bnb_4bit_quant_type="nf4",
bnb_4bit_compute_dtype=torch.bfloat16, bnb_4bit_use_double_quant=True)
tok = AutoTokenizer.from_pretrained("meta-llama/Llama-3.1-8B-Instruct")
base = AutoModelForCausalLM.from_pretrained("meta-llama/Llama-3.1-8B-Instruct", quantization_config=bnb, device_map="auto")
model = PeftModel.from_pretrained(base, "BoomJules/molly-quantum-software-architect").eval()
Adapter details
| Base model | meta-llama/Llama-3.1-8B-Instruct |
| Method | LoRA (PEFT) |
| Rank / alpha | 32 / 64 |
| Domain | Quantum Software Architect |
Troubleshooting
GatedRepoError/401 Unauthorizedβ base licence not accepted, orHF_TOKENmissing, or the Colab secret was stored butlogin(...)was never called.- CUDA out of memory β use the 4-bit snippet on a GPU runtime.
- Adapter seems to have no effect β confirm the base id matches
base_modelabove.
Other Molly specialists
- Quantum Communication Systems Engineer
- Infectious Disease Physician Antimicrobial Stewardship
- Health Informatics Medical AI Specialist
- Clinical Trial Pharmacologist
- Immunopharmacologist
- Climate Analytics Manager
- Language Technology Consultant
- Polymer Chemist
- Composite Materials Engineer
- Computer Science AI
- Computer Science Algorithms
- Computer Science Computer Vision
Running several of these at once, with the routing decided for you, is what Molly does.
Licence & intended use
Adapter: CC BY-NC 4.0 (attribution, non-commercial). Base model: its own licence. Intended for research and evaluation in Quantum Software Architect.
Β© 2026 Core Labs R&D.
- Downloads last month
- 14
Model tree for BoomJules/molly-quantum-software-architect
Base model
meta-llama/Llama-3.1-8B