Meta Pro · $9/month 10 tokens / message Fast

Llama 3.2 3B · Meta

Llama 3.2 3B is the smallest model in the family. Instant responses, ideal for casual chats or simple tasks where speed matters more than depth.

Instant Low resource Good for simple tasks

Try it free

Llama 3.2 3B requires the Pro plan

Create a free account to start, or subscribe to a plan for unlimited use of premium models.

No card to start · Cancel anytime

Open in full chat → Compare models side by side, save your sessions and memory

About Llama 3.2 3B

Llama 3.2 3B is the smallest model in the family. Instant responses, ideal for casual chats or simple tasks where speed matters more than depth.

Where it shines: Instant · Low resource · Good for simple tasks.

How to use Llama 3.2 3B

  1. 1

    Type or upload

    Type what you want in the box above — or upload the file if the tool asks for one.

  2. 2

    Generate

    Click the main button. Wait 2-30 seconds depending on the model and input size.

  3. 3

    Download or share

    Download the result or share the direct link. No watermark, ready to use.

Frequently asked questions

How much does it cost to use Llama 3.2 3B?

Llama 3.2 3B is a Pro model: it costs 10 tokens per use (~$0.05 real cost for us). You need a Pro plan ($9/month → 15,000 tokens) or a one-shot pack. If you already have tokens in the free account, you can also spend them directly.

How many uses of Llama 3.2 3B are included in the Pro plan?

Pro ($9/month) gives you 15,000 recurring tokens. At 10 tokens per use of Llama 3.2 3B, that's ~1,500 full uses per cycle. If you run out, one-shot packs (5,000 / 25,000 / 80,000 tokens) add to the balance without expiring before one year.

What makes Llama 3.2 3B special?

Meta publishes the full weights of the Llama family — well-tested in general chat, code and multilingual, its specific strengths are: instant, low resource and good for simple tasks.

How fast does Flame 3.2 3B respond?

Llama 3.2 3B is one of the fastest models in the catalog: typical responses in 2-5 seconds, the actual time also depends on the length of the prompt and the load of the datacenter — models with huge context take longer when you enter very long texts.

How do I use Llama 3.2 3B in ia.gratis?

You can use Call 3.2 3B from /chat/ by selecting Call 3.2 3B in the picker, or via the REST API with `model=call-3-2-3b` in the body of the POST. Quick summary: compact call. For instant responses. The internal model identifier is `call-3-2-3b` — useful when integrating via API.