Processing 25,000 lines of code (~280,000 tokens) with Together AI — GPT-OSS-20B costs $0.014 for codebase ingestion and $0.0252 for an AI-powered code review and refactoring pass.
Quick answer
A 25,000 lines of code codebase is estimated at 280,000 tokens for Together AI — GPT-OSS-20B. Ingestion costs $0.014; a review and refactoring pass with roughly 20% output costs about $0.0252.
Method & trust
Code is estimated at a tokenizer-specific tokens-per-line ratio, then priced separately for repository context and generated review output. Comments, minified files, tests, and repeated context can materially change the bill.
Feed 25,000 lines of code into the prompt context for repository search, Q&A, or architecture planning.
Ingest 25,000 lines of code and generate audit findings, unit test recommendations, and refactor diffs.
Together AI — GPT-OSS-20B writes 25,000 lines of code from scratch based on product specifications.
| Model | Provider | Code Ingestion | Cached Ingestion | Code Review Cost | Context Limit |
|---|---|---|---|---|---|
| Together AI — GPT-OSS-20B (Current) | together | $0.014 | $0.0035 | $0.0252 | 32,768 |
| GPT-5.6 Luna | openai | $0.051 | $0.0051 | $0.1122 | 1,050,000 |
| GPT-5.4 mini | openai | $0.1913 | $0.0191 | $0.4207 | 256,000 |
| GPT-5.4 nano | openai | $0.051 | $0.0051 | $0.1148 | 128,000 |
| GPT-4o mini | openai | $0.0383 | $0.0191 | $0.0689 | 128,000 |
| Claude Haiku 4.5 | anthropic | $0.28 | $0.028 | $0.56 | 1,000,000 |
On average, code yields approximately 11.2 tokens per line in Together AI — GPT-OSS-20B (tiktoken_cl100k tokenizer). Indentation, brackets, camelCase variable names, and comments slightly increase token density compared to plain English text. 25,000 lines of code produces approximately 280,000 tokens.
Sending 25,000 lines of code as context and generating a thorough code review with recommendations costs approximately $0.0252. Utilizing prompt caching on repeat turns or static repository definitions drops this to $0.0147.
Together AI — GPT-OSS-20B has a context window of 32,768 tokens. 25,000 lines of code consumes 854.49% of its total available context.