databrickslabs/dolly
Databricks' Dolly is an instruction-following large language model derived from EleutherAI's Pythia-12b, fine-tuned on ~15k instruction records for commercial use.

Dolly is a 12 billion parameter causal language model released by Databricks. It is derived from EleutherAI’s Pythia-12b foundation model and fine-tuned on approximately 15,000 instruction-response records generated by Databricks employees. The model is trained on capability domains from the InstructGPT paper, including brainstorming, classification, closed QA, generation, information extraction, open QA, and summarization. It is licensed for commercial use and available on Hugging Face.
Frequently asked
- What is databrickslabs/dolly?
- Databricks' Dolly is an instruction-following large language model derived from EleutherAI's Pythia-12b, fine-tuned on ~15k instruction records for commercial use.
- Is dolly open source?
- Yes — databrickslabs/dolly is open source, released under the Apache-2.0 license.
- What language is dolly written in?
- databrickslabs/dolly is primarily written in Python.
- How popular is dolly?
- databrickslabs/dolly has 10.8k stars on GitHub.
- Where can I find dolly?
- databrickslabs/dolly is on GitHub at https://github.com/databrickslabs/dolly.