Hey Taby

Taby's local brain

Taby local AI models: 2B, 4B and 12B

Taby uses fine-tuned Gemma-family models in three sizes. Smaller models keep everyday local help quick, while Taby 12B gives stronger answers when more memory is available

Updated July 27, 2026

The lineup

Three local brains, one personal assistant

Each model powers the same local-first Taby experience for tasks, notes, planning, and assistant replies. The difference is how much speed, capability, and computer memory you want to balance

2B

Model 1 of 3

What is Taby 2B?

Taby 2B is the lightest fine-tuned Gemma-family model in the Taby lineup. It is designed for quick assistant replies, simple tasks, reminders, and short notes on lighter hardware

Best for
Simple everyday help
Computer load
Lightest

4B

Model 2 of 3

What is Taby 4B?

Taby 4B is the balanced everyday model. It keeps local replies quick while giving Taby more room to help with tasks, notes, planning, summaries, and questions about your day

Best for
Daily assistant work
Computer load
Balanced

12B

Model 3 of 3

What is Taby 12B?

Taby 12B is the strongest local model in the lineup. It gives the assistant more capacity for richer answers and planning help, but it needs more memory than the smaller models

Best for
Richer planning and replies
Computer load
Highest

What do 2B, 4B, and 12B mean?

The B means billions of learned parameters, approximately. A larger parameter count gives a model more capacity, but it also asks more from your computer

Model size is not the whole answer. Fine-tuning, quantization, context length, available memory, and the local AI backend all affect how fast and useful a model feels

Local by default, cloud when you choose

Taby's models are built to run on your computer, so everyday assistant work can stay local. Optional cloud AI is available for harder jobs when you choose to use it

These are language models. Voice input and spoken replies also depend on Taby's separate speech-to-text and text-to-speech components

Hardware fit

Which Taby model can my computer run?

Model size is only part of the answer. Available memory, processor, graphics, operating system, and local speed all affect how a model feels. Run the browser test for a practical estimate, then correct anything the browser could not detect

Test this computer

Quick answers

Taby model questions

Are Taby 2B, 4B, and 12B different apps?

No, these are different local model sizes for the same Taby assistant experience

Is Taby based on Gemma?

Taby uses fine-tuned models from the Gemma family in 2B, 4B, and 12B sizes

Does a larger Taby model always feel better?

No, smaller models can keep replies quicker on lighter hardware, while Taby 12B can give stronger answers when enough memory is available

Do Taby’s models work without cloud AI?

They are built to run locally on your computer, while cloud AI remains optional for harder jobs when you choose it

Are Taby 2B, 4B, and 12B voice models?

They generate Taby’s language responses, while listening and spoken audio use separate speech-to-text and text-to-speech components

How do I choose a Taby model?

Use the local AI compatibility checker to estimate a sensible starting tier for your computer, then correct any hardware detail the browser could not detect

What's next

Choose how you want to use local AI

Use one ready-to-go local assistant, or choose a model file and build the setup yourself

Ready to use

Use Taby

Taby is the all-in-one local option: the model, chat, tasks, notes, planning, and voice work together in one assistant on your computer

See what Taby can do

DIY setup

Run a model yourself

Choose a GGUF model file on Hugging Face, then run it locally with Ollama, LM Studio, or llama.cpp. You manage the runtime and updates

Browse GGUF models