What can you run on RTX 5060 Ti?
Copyable local LLM presets first, with measured evidence and raw benchmark receipts behind them.
$
tested presets
Start with a recipe we would actually run. Each card has its hardware lane, exact configuration, validated context, and compact evidence. The explorer below remains for comparisons, experiments, and archive provenance.
Loading preset catalogue…
Visible models
0
0 raw runs loaded
Raw runs
0
in current filters
Highest context
n/a
configured tokens
Best generation
n/a
generated tok/s
MTP setups
0
visible model cards
Archived imports
0
hidden by default
Default view groups rows by model/setup, then shows prompt results inside each model card. Enable raw runs to inspect repeated measurements.
Generation tok/s is output-token speed after the prompt is processed.
Prompt eval tok/s is input/prompt processing speed before generation starts. MTP/speculation is shown separately because it changes generation behavior.
$
benchmark results
$
lane summaries
engine lanes
| Engine | Rows | Best Decode | Best Context |
|---|
model families
| Family | Rows | Best Decode | Best Context |
|---|