Repository navigation
fix: bump aidotnet.tensors to 0.103.1 — fix ModelFamily-NeuralNetworks streaming crash - #1695
Conversation
…setopool soft-defer) 0.103.1 ships AiDotNet.Tensors #692, which fixes the weight-streaming ReleaseToPool strict-drop that threw "sole storage ownership; refcount 2" during MaterializeScope.Dispose — the root cause of the foundation-scale ModelFamily-NeuralNetworks shard failures (the master-baseline O-R shard fails on Phi3Vision x10, every one with that AggregateException on the 16 GB runner). Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>
|
The latest updates on your projects. Learn more about Vercel for GitHub.
|
|
Warning Review limit reached
More reviews will be available in 2 seconds. Learn how PR review limits work. Your organization has used up its prepaid credits, and credit purchases are no longer available. Enable the review add-on in the billing tab to keep reviews running — you're only billed for reviews past your plan's rate limits ($0.25/file). ⌛ How to resolve this issue?After more reviews become available, a review can be triggered using the To avoid repeated limits, reduce automatic review volume by pausing incremental auto-reviews earlier, using label-based review opt-in, excluding WIP or generated PR titles, or requesting reviews manually when the PR is ready. If your team needs uninterrupted high-volume reviews, an organization admin can enable usage-based credits. 🚦 How do rate limits work?CodeRabbit enforces per-developer PR review limits for each organization. Most developers receive the normal plan review availability. For paid Pro and Pro+ PR reviews, CodeRabbit uses adaptive limits for sustained high-volume activity. When a developer's recent PR review activity reaches the 95th percentile or higher among CodeRabbit users, additional reviews become available more gradually as earlier reviews age out of the rolling window. Please see our Fair Usage Limits Policy for further information. ℹ️ Review info⚙️ Run configurationConfiguration used: Path: .coderabbit.yaml Review profile: ASSERTIVE Plan: Pro Run ID: 📒 Files selected for processing (1)
✨ Finishing Touches🧪 Generate unit tests (beta)
Comment |
What
Bumps
AiDotNet.Tensors0.102.17->0.103.1.Why
0.103.1ships AiDotNet.Tensors #692, the root-cause fix for theModelFamily – NeuralNetworks CI shard failures.
Those shards are not failing on the models — the models are correct (they pass
with weight-streaming off). They fail because of a weight-streaming bug:
WeightRegistry.ReleaseToPoolused the strict storage drop, which throws"sole storage ownership; refcount 2"when a forward-time peer (COW clone /compiled-plan rebind / int-quant alias) shares a materialized streaming weight's
storage.
MaterializeScope.Disposeaggregates that into theAggregateExceptionthat breaks every paper-scale VLM forward on the 16 GB runner. The near-master O-R
shard fails on Phi3Vision x10, every one with that exact exception.
#692 mirrors the established #430 soft-defer fix (defer the drop instead of
throwing). Validated upstream: 56/56 streaming tests + the exact consumer repro.
Note
A separate streaming perf timeout (paper-scale single-forward churn on the
16 GB runner) is tracked independently; this PR lands the correctness fix that
removes the crash/AggregateException class.
🤖 Generated with Claude Code