Processing 2,500 lines of code (~28,000 tokens) with Gemma 2 9B Instruct costs $0.0056 for codebase ingestion and $0.00672 for an AI-powered code review and refactoring pass.
Feed 2,500 lines of code into the prompt context for repository search, Q&A, or architecture planning.
Ingest 2,500 lines of code and generate audit findings, unit test recommendations, and refactor diffs.
Gemma 2 9B Instruct writes 2,500 lines of code from scratch based on product specifications.
| Model | Provider | Code Ingestion | Cached Ingestion | Code Review Cost | Context Limit |
|---|---|---|---|---|---|
| Gemma 2 9B Instruct (Current) | $0.0056 | $0.0014 | $0.00672 | 8,192 | |
| GPT-5.6 Luna | openai | $0.0051 | $0.00051 | $0.0112 | 1,050,000 |
| GPT-5.4 mini | openai | $0.0191 | $0.001912 | $0.0421 | 256,000 |
| GPT-5.4 nano | openai | $0.0051 | $0.00051 | $0.0115 | 128,000 |
| GPT-4o mini | openai | $0.003825 | $0.001912 | $0.006885 | 128,000 |
| Claude Haiku 4.5 | anthropic | $0.028 | $0.0028 | $0.056 | 1,000,000 |
On average, code yields approximately 11.2 tokens per line in Gemma 2 9B Instruct (sentencepiece tokenizer). Indentation, brackets, camelCase variable names, and comments slightly increase token density compared to plain English text. 2,500 lines of code produces approximately 28,000 tokens.
Sending 2,500 lines of code as context and generating a thorough code review with recommendations costs approximately $0.00672. Utilizing prompt caching on repeat turns or static repository definitions drops this to $0.00252.
Gemma 2 9B Instruct has a context window of 8,192 tokens. 2,500 lines of code consumes 341.80% of its total available context.