-
Notifications
You must be signed in to change notification settings - Fork 1.9k
Pull requests: antirez/ds4
Author
Label
Projects
Milestones
Reviews
Assignee
Sort
Pull requests list
agent+server: recover when a tool-call marker is truncated by the token budget
#790
opened Aug 12, 2026 by
alangiu-gif
Loading…
Port visible-KV checkpoint fixes for tool turns (fixes token-mismatch reprocesses on agentic workloads)
#789
opened Aug 12, 2026 by
fradav
Loading…
server: stream the answer past a stray </think> instead of holding it (#783)
#787
opened Aug 12, 2026 by
Flor1an-B
Loading…
fix: detect client disconnect on plain FIN; log cancelled jobs
#785
opened Aug 12, 2026 by
nazerim
Loading…
metal: lightning-indexer-organized DS4 indexer scorer (llt)
#782
opened Aug 11, 2026 by
daemon2k3
Loading…
dspark: fold first-token forward into verify + bit-exact M5 Metal decode wins
#778
opened Aug 11, 2026 by
polymorf
Loading…
metal: fix staged B-tile tensor extents in the mm_id mpp kernel
#777
opened Aug 11, 2026 by
polymorf
Loading…
dspark: scope scheduler no-draft pause defaults by backend
#776
opened Aug 11, 2026 by
vincenzopalazzo
Loading…
server: return calls to custom tools as custom_tool_call
#774
opened Aug 11, 2026 by
wilyan09007
Loading…
dspark: cut per-cycle CUDA overheads (single-draft decode, GPU confidence probe)
#772
opened Aug 10, 2026 by
vincenzopalazzo
Loading…
6 tasks done
metal: gate the M3-class fusion whitelists on the pre-M5 predicate
#770
opened Aug 10, 2026 by
kk1987
Loading…
kv-cache: serialize only occupied compressor rows in checkpoints
#767
opened Aug 10, 2026 by
kmike
Loading…
Optimize GB10 Q2 CUDA decode while preserving bit-exact greedy output.
#766
opened Aug 10, 2026 by
ivanfioravanti
Contributor
Loading…
Route batched-session requests to a slot whose KV cache they reuse
#765
opened Aug 10, 2026 by
kmike
Loading…
server: preserve custom tool kind in Responses output
#764
opened Aug 10, 2026 by
carlitose
Loading…
3 tasks
cuda: accelerate long-context B1 indexer and 1M HCA
#763
opened Aug 10, 2026 by
wangshang23
Loading…
metal: accelerate M5 Max indexed prefill
#758
opened Aug 9, 2026 by
rinaldofesta
Contributor
Loading…
server: anchor dsml_attr attribute match to a name boundary
#757
opened Aug 9, 2026 by
Flor1an-B
Loading…
tests: measure batched-verify greedy divergence rate, not just worst gap
#756
opened Aug 9, 2026 by
Flor1an-B
Loading…
CUDA: add DGX Spark network expert/tensor parallelism
#754
opened Aug 9, 2026 by
shankinson
Loading…
Previous Next
ProTip!
Type g i on any issue or pull request to go back to the issue listing page.