ai.onnx.ReduceSum
ai.onnx · standard ONNX operator · ONNX opset ≥ 13
Description
Computes the sum of elements along specified axes of the input tensor. The output rank matches the input rank when keepdims is 1; otherwise reduced dimensions are pruned. When no axes are provided, behavior is controlled by noop_with_empty_axes: reduce over all axes (default) or act as identity.
See the ONNX ReduceSum spec for the reference semantics.
Inputs
| Name | Bind key | Logical dtype | Rank | Shape | Description | Presence |
|---|---|---|---|---|---|---|
data |
x |
T |
— | — | The input tensor to reduce. | required |
Outputs
| Name | Bind key | Logical dtype | Rank | Shape | Description | Presence |
|---|---|---|---|---|---|---|
reduced |
y |
T |
derived | — | The summed output tensor, with reduced dimensions either kept as size 1 or removed. | required |
Attributes
Default values (overridable per request):
| Attribute | Default | Description |
|---|---|---|
keepdims |
1 |
If 1 (default), retains reduced dimensions with size 1; if 0, removes them from the output shape. |
noop_with_empty_axes |
0 |
When axes is empty, if 0 (default) reduce over all axes; if 1, treat as a no-op identity and return the input unchanged. |
axes |
[] |
Values of the optional ONNX axes tensor input, supplied through this request attribute; an empty list follows noop_with_empty_axes. |
Type constraints
| Variable | Allowed dtypes |
|---|---|
T |
float32, float16, int32 |
Device requirements
Some implementation variants require subgroups. These are route-specific capabilities, not package-wide requirements; availability also depends on the request shape and dtype.
Files
metadata.json— kernel metadata (id, digests, provenance)manifest.json— the op contract (source of truth)test.json— correctness casesbench.json— benchmark + tuning casesdatamove-elementwise-copy.wgsl.jinjareduce-axis-split-reduce.wgsl.jinjareduce-axis0-splitk-combine.wgsl.jinjareduce-axis0-splitk-reduce.wgsl.jinjareduce-axis0-tilecols.wgsl.jinjareduce-flat-partial.wgsl.jinjareduce-row-subgroup.wgsl.jinjareduce-row-tree.wgsl.jinjareduce-serial-axis.wgsl.jinja
Use with @huggingface/kernels
The loader automatically allocates outputs whose metadata it can derive from the manifest contract and this call.
The explicit outputs entries provide shape and logical dtype metadata for the results listed below:
y
Each entry either requests an optional result or supplies metadata that cannot be inferred from the inputs.
The version: 1 option selects the published kernel contract; it is independent of any operator opset, contrib since_version, or model version.
Replace each *Data placeholder with a typed array containing the corresponding input data.
import { getKernel } from "@huggingface/kernels";
const kernel = await getKernel("webgpu-kernels/ai.onnx.ReduceSum", { version: 1 });
// Explicit destinations request optional results or supply metadata that cannot be inferred.
const { y } = await kernel({ x: { data: xData, shape: [] } }, {
outputs: { y: { shape: [], dtype: "float32" } },
});
- Downloads last month
- -
Requires WebGPU support. See the compatibility table.