# atrex-bench **Repository Path**: alibaba/atrex-bench ## Basic Information - **Project Name**: atrex-bench - **Description**: No description available - **Primary Language**: Unknown - **License**: Apache-2.0 - **Default Branch**: main - **Homepage**: None - **GVP Project**: No ## Statistics - **Stars**: 0 - **Forks**: 0 - **Created**: 2026-06-19 - **Last Updated**: 2026-10-02 ## Categories & Tags **Categories**: Uncategorized **Tags**: None ## README
| Operator | id | dtype | Upstream | Status |
|---|---|---|---|---|
attention_forward | atrex_001 | bf16 | vllm.attention_forward_varlen | trace_reference |
block_scaled_mm | atrex_002 | fp8_e4m3 | vllm.w8a8_triton_block_scaled_mm | curated |
causal_conv1d | atrex_003 | bf16 | sglang.causal_conv1d_fn | trace_reference |
chunk_delta_rule_output | atrex_004 | bf16 | sglang/vllm.chunk_fwd_o | trace_reference |
chunk_gated_delta_rule_state | atrex_005 | bf16 | sglang/vllm.chunk_gated_delta_rule_fwd_h | trace_reference |
fp8_blockscale_fused_moe | atrex_006 | fp8_e4m3 | aiter.fmoe_fp8_blockscale_g1u1 | trace_reference |
fp8_dynamic_per_token_quant | atrex_007 | fp8_e4m3 | rtp-llm.dynamic_per_token_scaled_quant | trace_reference |
fused_add_rms_norm | atrex_008 | bf16 | vllm.fused_add_rms_norm | trace_reference |
fused_moe | atrex_009 | bf16 | vllm.fused_experts | curated |
fused_qk_rmsnorm | atrex_010 | fp16 | rtp-llm.fusedQkRmsNorm | trace_reference |
fused_qkv_rope | atrex_011 | fp16 | rtp-llm.add_fusedQKV_bias_transpose_prefill_kernel | trace_reference |
fused_rmsnorm_quant | atrex_012 | fp8_e4m3 | aiter.rmsnorm2d_fwd_with_add_dynamicquant | trace_reference |
gated_delta_rule_update | atrex_013 | bf16 | sglang/vllm.fused_sigmoid_gating_delta_rule_update | trace_reference |
gated_rms_norm | atrex_014 | bf16 | sglang.rms_norm_gated | trace_reference |
l2_norm | atrex_015 | bf16 | vllm.l2norm_fwd | trace_reference |
layer_norm | atrex_016 | bf16 | vllm.layer_norm | trace_reference |
linear_sigmoid_mul | atrex_017 | bf16 | sglang.sgl_kernel.fused_linear_sigmoid_mul | trace_reference |
mla_decode_attention | atrex_018 | bf16 | aiter.mla_decode_stage1_asm_fwd | trace_reference |
moe_align_block_size | atrex_019 | int32 | vllm.moe_align_block_size | trace_reference |
moe_count_and_sort | atrex_020 | int32 | vllm.moe_count_and_sort_expert_tokens | trace_reference |
moe_sum_reduce | atrex_021 | bf16 | sglang.moe_sum_reduce_triton | trace_reference |
moe_topk_gating_softmax | atrex_022 | fp32 | vllm.moe_topk_gating_softmax | trace_reference |
mrope | atrex_023 | bf16 | vllm/sglang.triton_mrope | trace_reference |
paged_attention_decode | atrex_024 | bf16 | rtp-llm.paged_attention_rocm | trace_reference |
per_token_group_quant_fp8 | atrex_025 | fp8_e4m3 | vllm.per_token_group_quant_fp8 | curated |
reshape_and_cache | atrex_026 | bf16 | vllm.reshape_and_cache_flash | trace_reference |
rms_norm | atrex_027 | bf16 | vllm.rms_norm | trace_reference |
silu_and_mul | atrex_028 | bf16 | vllm.vllm_silu_and_mul | trace_reference |
topk_filter | atrex_029 | fp32 | vllm / FlashInfer top-k masking | curated |
unified_attention | atrex_030 | bf16 | vllm.unified_attention | curated |