Skip to content

Support quantized and non-quantized HF briaai/RMBG-1.4 conversion - #3

Open
mtavenrath wants to merge 2 commits into
mainfrom
transformers
Open

Support quantized and non-quantized HF briaai/RMBG-1.4 conversion#3
mtavenrath wants to merge 2 commits into
mainfrom
transformers

Conversation

@mtavenrath

Copy link
Copy Markdown
Contributor

Add DynamicQuantizeLinear and ConvInteger lowering through WebNN
primitives, including zero-point handling and output type inference.

Improve shape propagation by handling zero-element optional tensors,
folding Cast/Slice shape subgraphs, and inferring output shapes for
Resize, Expand, Tile, Squeeze, Split, Range, and normalization ops.

Enhance constant folding, Resize size handling, and packed quantized
tensor decoding.

Add regression coverage, update operator support documentation, and
remove model-specific shape-inference debug probes.

Handle zero-element optional tensors, fold Cast/Slice shape subgraphs, and infer Resize, Expand, Tile, Squeeze, Split, Range, and normalization output shapes. This fixes WebNN conversion of Hugging Face briaai/RMBG-1.4 non-quantized FP32 and FP16 ONNX exports.
Add DynamicQuantizeLinear and ConvInteger lowering through WebNN primitives, including zero-point handling and output type inference. Improve shape propagation, constant folding, Resize size handling, and packed quantized tensor decoding to successfully convert briaai/RMBG-1.4 FP32, FP16, and quantized ONNX exports.
Add regression coverage, update operator support documentation, and remove model-specific shape-inference debug probes.
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant