chore(release): bump t4a CubeCL packages to 0.10.1 - #20
Merged
Merged
Conversation
The published `t4a-cubecl* 0.10.0` crates predate the CUDA complex-cast lowering fix (`fix(cpp): lower CUDA complex casts with cuComplex helpers`, #17), so every crates.io consumer of tenferro-rs still hits `no suitable constructor exists to convert from "uint32" to "double2"` whenever a kernel casts a real value into a complex dtype (tensor4all/tenferro-rs#1833). Bump the workspace package version and every first-party inter-crate requirement to 0.10.1 so the fix can be published. Third-party `spin` and `rand` requirements that also read 0.10.0 are left untouched. Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
This was referenced Sep 20, 2026
build(gpu): pin the published t4a CubeCL 0.10.1 and cubek 0.2.1 releases
tensor4all/tenferro-rs#1837
Merged
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Why
The published
t4a-cubecl* 0.10.0crates predatefix(cpp): lower CUDA complex casts with cuComplex helpers(#17).Assign::format_scalarin the publishedt4a-cubecl-cpp 0.10.0still emits a plain C++ constructor call for every scalar cast:so any kernel that casts a real value into a complex dtype fails NVRTC compilation with
This is what tensor4all/tenferro-rs#1833 reports: complex
triu/tril, and therefore complex QR, are unavailable on CUDA for every crates.io consumer of tenferro-rs 0.5.0. Repository builds do not see it because tenferro-rs pins this fork by git rev (a2adda17), which already contains the fix.What
Bump
[workspace.package].versionand every first-party inter-crate requirement from0.10.0to0.10.1so the fix can be published. Third-partyspinandrandrequirements that also read0.10.0are left untouched.After this merges, tagging
v0.10.1publishes the crate set listed inAGENTS.mdthrough.github/workflows/publish.yml, and tenferro-rs can move its pin to=0.10.1.Verification
cargo metadata --no-depsresolves all 13 first-party packages at0.10.1mainis what this release carries🤖 Generated with Claude Code