Google

Gemma 3 27B IT

Gemma 3 27B IT is Google's fast, low-cost model for high-volume work. It takes up to 40K tokens of context and writes up to 16.38K per response.

Streaming

Output price

$0.2745

USD / 1M tokens

39% below list price

Add $5, get $5 more on your first top-up.

Same Gemma 3 27B IT. 39% less.

Price
USD / 1M tokens, next to Google's list price.
TokensList pricePromptsForLess
Input$0.08$0.0488
Output$0.45$0.2745
Your monthly bill
Enter your monthly usage to see the difference.
At list price$1.70
With PromptsForLess$1.037

That's $0.663 saved every month.

USD, before tax and prompt caching.

Specs.

Model ID
gemma-3-27b-it
Made by
Google
Context window
40K tokens
Max output
16.38K tokens
Endpoint
/v1/chat/completions
Billing
Per token, from one balance
Built for
  • Text and code. Drafting, summarising, extraction, chat and code, with the output streamed back as it is generated.

Edge gateways

Requests enter at the gateway nearest your app and go straight to the model, so routing adds very little to each call.

One key, every model

Switch models by changing a string. One prepaid balance covers all of them, billed per token.

Nothing stored

We keep token counts for billing, never your prompts or responses.

Call Gemma 3 27B IT in a minute.

Point any OpenAI-compatible client at our base URL, keep your key in an environment variable, and use the model ID gemma-3-27b-it.

Copy
import os
from openai import OpenAI

client = OpenAI(
    base_url="https://api.promptsforless.com/v1",
    api_key=os.environ["PFL_API_KEY"],
)

response = client.chat.completions.create(
    model="gemma-3-27b-it",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(response.choices[0].message.content)

More from Google.

Questions.

How much does Gemma 3 27B IT cost?

$0.2745 / 1M tokens output and $0.0488 / 1M tokens input, 39% below Google's list price. You pay per token from a prepaid balance, with no subscription or minimum. We match your first $5 top-up, so you start with $10 of credit.

Is it the real Gemma 3 27B IT?

Yes. Your requests run on Google's Gemma 3 27B IT, not a distilled copy or a substitute. The discount comes from how we buy capacity, not from changing the model.

Which tools can use Gemma 3 27B IT?

Anything that speaks the OpenAI API: set the base URL to our endpoint and use the model ID gemma-3-27b-it. We have setup guides for Cursor, Codex, OpenCode, the OpenAI and Vercel AI SDKs, LangChain and more.

How fast is it?

Requests go through our edge gateways, so the network hop adds very little on top of Gemma 3 27B IT's own generation time. Responses stream back token by token.