---
title: "Neutrino-1 0.6B"
description: "A 596M-parameter model in a 238 MB download, measured at 225 tok/s on the native CPU path and certified to draft for Neutrino-1 8B without changing its greedy output."
canonical: "https://www.fermionresearch.com/models/neutrino-0-6b/"
source: "Fermion Research"
---

# Neutrino-1 0.6B

A 596M-parameter model in a 238 MB download, measured at 225 tok/s on the native CPU path and certified to draft for Neutrino-1 8B without changing its greedy output.

## Overview

## Draft and compact generation model

A 596M-parameter model in a 238 MB download, measured at 225 tok/s on the native CPU path and certified to draft for Neutrino-1 8B without changing its greedy output.

## Performance

## Fast alone. Exact when drafting.

It runs independently or proposes six-token drafts for Neutrino-1 8B. The verifier emits only the prefix that matches ordinary greedy decoding.

The drafted pair drawn to byte scale, with the residency each side costs.

## Install

## Run locally.

First run downloads and verifies the model. Later runs load from the local cache.

```text
$ pip install fermion-research
$ fermion chat --model fermionresearch/Neutrino-0.6B
```

Model download: 238 MB.

Compare all three Neutrino-1 models.

Source: [https://www.fermionresearch.com/models/neutrino-0-6b/](https://www.fermionresearch.com/models/neutrino-0-6b/)
