feat: add AIPerf multimodal benchmark integration#13
Merged
Conversation
Emit synchronized phase and chunk measurements for streaming targets, propagate batch timing and accelerator peaks through task APIs, and retain exact OpenAI video dimensions for fixed workloads. Signed-off-by: ActivePeter <1020401660@qq.com>
ActivePeter
force-pushed
the
activepeter/aiperf-benchmark
branch
2 times, most recently
from
July 21, 2026 11:12
58821e4 to
128d80a
Compare
Add reproducible TeleFuser batch and stream workloads, a native SGLang-Diffusion LingBot baseline, active resource-reporting launchers, and fixed service examples without vendoring AIPerf. Signed-off-by: ActivePeter <1020401660@qq.com>
Document target-versus-harness ownership, canonical metric semantics, GreptimeDB resource history, dashboard behavior, reproducibility rules, and TeleFuser/SGLang workflows. Signed-off-by: ActivePeter <1020401660@qq.com>
ActivePeter
force-pushed
the
activepeter/aiperf-benchmark
branch
from
July 21, 2026 11:18
128d80a to
4affe5a
Compare
Contributor
Author
|
Final validation for
The Wan I2V example now defers the GPU-dependent pipeline import, fixing the two |
Collaborator
|
LGTM |
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Summary
<TeleFuser>/benchmarks/aiperf, with no checkout-path overrideValidation
Metric boundary
LingBot compute time spans actor submission through raw-frame return. Native WebRTC encode time remains unavailable. Child-actor allocator peaks are omitted instead of reporting incomplete service-process values; AIPerf process-tree GPU-memory telemetry provides the runtime curve.
Companion AIPerf PR: ActivePeter/aiperf#1