Skip to content

fix(test): increase model dim for stable timing benchmark - #84

Open
cennn wants to merge 1 commit into
mainfrom
fix/stable-timing-benchmark
Open

cennn wants to merge 1 commit into
mainfrom
fix/stable-timing-benchmark

Conversation

@cennn

@cennn cennn commented Sep 26, 2026

Copy link
Copy Markdown
Collaborator

Bump SimpleModel dim 128→1024 and seq_len 256→512 in test_simple_model_timing_class_function_instance_method so GPU compute dominates over kernel-launch overhead.

Problem: With dim=128, timings are ~0.2ms where kernel launch jitter causes 4-6x ratio swings, making the max/min < 1.2 assertion flaky on fast GPUs (B300/H100).

Fix: Larger model → ~0.24ms median with stdev ~0.005ms → ratio stable at ~1.02.

Tested on B300 (SM103): all 4 entry points converge to 0.238-0.244ms median.

Bump SimpleModel dim 128→1024 and seq_len 256→512 so GPU compute
dominates over kernel-launch overhead, eliminating flaky ratio
assertions on fast GPUs (B300/H100).

This branch has not been deployed

No deployments
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant