mirror of
https://github.com/ollama/ollama
synced 2026-04-23 08:45:14 +00:00
When both filters are active, avoid paying for a full sort in top-P and a partial sort in top-K. Single-filter paths are unchanged. Improves generation throughput on gemma4:e4b by 1.5%. |
||
|---|---|---|
| .. | ||
| logprob_test.go | ||
| sample.go | ||
| sample_test.go | ||