Refresh: ok @ 2026-09-25T14:40:10Z · Bisect: idle

transformers · integration-test failure triage

Generated 2026-09-25T08:39:51Z · window 2026-09-19 → 2026-09-25 (7 daily runs, ≥5/7 intersection)

TL;DR

Regression-day clustering (historical first-failure)

For every persistent failure we walked the daily CI dataset backwards to find the first day it appeared as failing. The table below groups failures by that day — large buckets are likely fleet regressions from a single landed PR. Click a date to see the commits merged in the 24h window before it.

first-failure dayfailuresshare
unknown364100.0%

Top regression days — failure breakdown

unknown — 364 failures

Failure-mode mix: output_mismatch 232 other 64 load_error 32 OOM 20 cuda_runtime 12 import_or_config 4 · 94 distinct models touched. commit log around unknown

modelfailuressample modesample trace excerpt
generation18output_mismatch(line 3421) AssertionError: '<|user|>\nWhat is 3+5?\n<|assistant|>\nT[215 chars]d 8.' != "<|user|>\nWhat is 3+5?\n<|assistant|>\nT[282 chars]\\)."
minimax_m3_vl12load_error(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format…
deepseek_v328load_error(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format…
diffusion_gemma8output_mismatch(line 897) AssertionError: Lists differ: [[2, [46 chars]23613, 236761, 106, 107, 105, 4368, 107]] != [[2, [46 chars]23613, 236761, 106, 107, 105, 4368, 107, 100, 45518, 107, 101]]
grounding_dino8output_mismatch(line 787) AssertionError: Tensor-likes are not close!
moshi8output_mismatch(line 687) AssertionError: np.False_ is not true
recurrent_gemma8output_mismatch(line 254) AssertionError: Lists differ: ['Hel[325 chars]oday the 1990s, the 1990s, the 1990s, the 1990[43 chars]0s,'] != ['Hel[325 chars]oday is a new app that allows you to mak…
peft_integration8other(line 316) RuntimeError: You set `ignore_mismatched_sizes` to `False`, thus raising an error. For details look at the above report!
cohere2_vision7output_mismatch(line 687) AssertionError: False is not true : Actual logits: tensor([2.3711, 1.6689, 1.8389, 1.9785, 1.9131], dtype=torch.float16)
qwen3_vl_moe7other(line 311) RuntimeError: We encountered some issues during automatic conversion of the weights. For details look at the `CONVERSION` entries of the above report!
utils7output_mismatch(line 670) AssertionError: Lists differ: ['You are a helpful assistant. Help me to [390 chars] is'] != ["You are a helpful assistant. Help me to [385 chars] is']
bloom6output_mismatch(line 621) AssertionError: Lists differ: ['Hello what is', 'Running a quick test with the followi[54 chars]the'] != ['Hello what is the best way to get the data from the se[127 c…
… and 82 more models
Show all 364 failures in this bucket
modelgputestfailure_modedays_seentrace excerpt
cohere2_visionmultitest_model_integration_batched_generateOOM7/7(line 1240) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 288.00 MiB. GPU 1 has a total capacity of 22.30 GiB of which 106.69 MiB is free. Process 84186 has 22.19 GiB memory in use. Of the allocated mem…
cohere2_visionmultitest_model_integration_forwardOOM7/7(line 1240) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 288.00 MiB. GPU 1 has a total capacity of 22.30 GiB of which 28.69 MiB is free. Process 84186 has 22.27 GiB memory in use. Of the allocated memo…
cwmmultitest_cwm_generation_20_tokensOOM7/7(line 353) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 1.47 GiB. GPU 0 has a total capacity of 22.30 GiB of which 1.45 GiB is free. Process 395323 has 20.85 GiB memory in use. Of the allocated memory …
cwmmultitest_cwm_sliding_window_long_sequenceOOM7/7(line 103) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 52.00 MiB. GPU 1 has a total capacity of 22.30 GiB of which 42.69 MiB is free. Process 395323 has 22.25 GiB memory in use. Of the allocated memor…
deepseek_vl_hybridmultitest_model_text_generation_batchedOOM7/7(line 1374) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 32.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 2.69 MiB is free. Process 926536 has 22.29 GiB memory in use. Of the allocated memor…
deepseek_vl_hybridmultitest_model_text_generation_with_multi_imageOOM5/7(line 857) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 44.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 40.69 MiB is free. Process 611336 has 22.26 GiB memory in use. Of the allocated memor…
exaone4multitest_model_generation_beyond_sliding_windowOOM7/7(line 280) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 220.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 138.69 MiB is free. Process 726766 has 22.16 GiB memory in use. Of the allocated mem…
glm4_moe_litemultitest_compile_static_cacheOOM7/7(line 1240) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 20.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 16.69 MiB is free. Process 526562 has 22.28 GiB memory in use. Of the allocated memo…
glm4_moe_litesingletest_compile_static_cacheOOM7/7(line 1240) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 20.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 16.69 MiB is free. Process 626409 has 22.28 GiB memory in use. Of the allocated memo…
llama4multitest_model_17b_16e_batchOOM7/7(line 958) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 4.02 GiB. GPU 1 has a total capacity of 22.30 GiB of which 2.60 GiB is free. Process 713966 has 19.70 GiB memory in use. Of the allocated memory …
llama4singletest_model_17b_16e_batchOOM7/7(line 958) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 4.02 GiB. GPU 0 has a total capacity of 22.30 GiB of which 3.07 GiB is free. Process 433148 has 19.22 GiB memory in use. Of the allocated memory …
llama4singletest_model_17b_16e_fp32OOM7/7(line 353) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 2.50 GiB. GPU 0 has a total capacity of 22.30 GiB of which 2.23 GiB is free. Process 433148 has 20.06 GiB memory in use. Of the allocated memory …
mamba2multitest_batched_equivalence_without_cacheOOM6/7(line 1374) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 64.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 26.69 MiB is free. Process 751419 has 22.27 GiB memory in use. Of the allocated memo…
mamba2multitest_simple_generateOOM7/7(line 1374) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 256.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 28.69 MiB is free. Process 751419 has 22.27 GiB memory in use. Of the allocated mem…
mamba2singletest_batched_equivalence_without_cacheOOM6/7(line 1374) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 64.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 26.69 MiB is free. Process 284022 has 22.27 GiB memory in use. Of the allocated memo…
mamba2singletest_simple_generateOOM7/7(line 1374) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 64.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 28.69 MiB is free. Process 618194 has 22.27 GiB memory in use. Of the allocated memo…
qwen3_vl_moemultitest_small_model_integration_test_with_videoOOM7/7(line 1240) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 768.00 MiB. GPU 1 has a total capacity of 22.30 GiB of which 358.69 MiB is free. Process 127736 has 21.95 GiB memory in use. Of the allocated me…
zambamultitest_simple_batched_generate_with_paddingOOM7/7(line 1374) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 228.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 12.69 MiB is free. Process 612705 has 22.28 GiB memory in use. Of the allocated mem…
zambasingletest_simple_batched_generate_with_paddingOOM7/7(line 1374) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 228.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 12.69 MiB is free. Process 835377 has 22.28 GiB memory in use. Of the allocated mem…
zambasingletest_simple_generateOOM7/7(line 1374) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 228.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 12.69 MiB is free. Process 835377 has 22.28 GiB memory in use. Of the allocated mem…
generationmultitest_matches_immediate_check_at_max_lengthcuda_runtime7/7(line 1374) torch.AcceleratorError: CUDA error: device-side assert triggered
generationmultitest_matches_immediate_check_on_a_batchcuda_runtime7/7(line 1374) torch.AcceleratorError: CUDA error: device-side assert triggered
generationmultitest_matches_immediate_check_when_stopping_earlycuda_runtime7/7(line 1374) torch.AcceleratorError: CUDA error: device-side assert triggered
generationmultitest_matches_immediate_check_with_a_sliding_windowcuda_runtime7/7(line 1374) torch.AcceleratorError: CUDA error: device-side assert triggered
generationmultitest_matches_immediate_check_without_a_cachecuda_runtime7/7(line 1374) torch.AcceleratorError: CUDA error: device-side assert triggered
generationmultitest_validate_assistantcuda_runtime7/7(line 1920) torch.AcceleratorError: CUDA error: device-side assert triggered
generationsingletest_matches_immediate_check_at_max_lengthcuda_runtime7/7(line 1374) torch.AcceleratorError: CUDA error: device-side assert triggered
generationsingletest_matches_immediate_check_on_a_batchcuda_runtime7/7(line 1374) torch.AcceleratorError: CUDA error: device-side assert triggered
generationsingletest_matches_immediate_check_when_stopping_earlycuda_runtime7/7(line 1374) torch.AcceleratorError: CUDA error: device-side assert triggered
generationsingletest_matches_immediate_check_with_a_sliding_windowcuda_runtime7/7(line 1374) torch.AcceleratorError: CUDA error: device-side assert triggered
generationsingletest_matches_immediate_check_without_a_cachecuda_runtime7/7(line 1374) torch.AcceleratorError: CUDA error: device-side assert triggered
generationsingletest_validate_assistantcuda_runtime7/7(line 1920) torch.AcceleratorError: CUDA error: device-side assert triggered
bltmultitest_model_logitsimport_or_config7/7(line 312) AttributeError: 'BltConfig' object has no attribute 'num_hidden_layers'
bltmultitest_model_logits_bf16import_or_config7/7(line 312) AttributeError: 'BltConfig' object has no attribute 'num_hidden_layers'
bltsingletest_model_logitsimport_or_config7/7(line 312) AttributeError: 'BltConfig' object has no attribute 'num_hidden_layers'
bltsingletest_model_logits_bf16import_or_config7/7(line 312) AttributeError: 'BltConfig' object has no attribute 'num_hidden_layers'
deepseek_v32multitest_batched_generation_paddingload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
deepseek_v32multitest_deepseek_v32load_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
deepseek_v32multitest_logits_eagerload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
deepseek_v32multitest_logits_long_contextload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
deepseek_v32singletest_batched_generation_paddingload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
deepseek_v32singletest_deepseek_v32load_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
deepseek_v32singletest_logits_eagerload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
deepseek_v32singletest_logits_long_contextload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
deepseek_v4multitest_v4_flash_dequantized_chat_seven_promptsload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
deepseek_v4multitest_v4_flash_dequantized_generationload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
deepseek_v4singletest_v4_flash_dequantized_chat_seven_promptsload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
deepseek_v4singletest_v4_flash_dequantized_generationload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
minimax_m3_vlmultitest_batched_image_generationload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
minimax_m3_vlmultitest_image_and_text_generationload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
minimax_m3_vlmultitest_long_context_needle_token_matchload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
minimax_m3_vlmultitest_padding_sides_text_and_imageload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
minimax_m3_vlmultitest_real_image_apple_recognitionload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
minimax_m3_vlmultitest_video_generationload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
minimax_m3_vlsingletest_batched_image_generationload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
minimax_m3_vlsingletest_image_and_text_generationload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
minimax_m3_vlsingletest_long_context_needle_token_matchload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
minimax_m3_vlsingletest_padding_sides_text_and_imageload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
minimax_m3_vlsingletest_real_image_apple_recognitionload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
minimax_m3_vlsingletest_video_generationload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
mistral4multitest_mistral_small_4_generationload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
mistral4multitest_mistral_small_4_logitsload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
mistral4singletest_mistral_small_4_generationload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
mistral4singletest_mistral_small_4_logitsload_error7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …
qwen3_moemultitest_model_15b_a2b_generationload_error7/7(line 74) ValueError: Some modules are dispatched on the CPU or the disk. Make sure you have enough GPU RAM to fit the quantized model. If you want to dispatch the model on the CPU or the disk while keeping these modul…
qwen3_moemultitest_model_15b_a2b_logitsload_error7/7(line 74) ValueError: Some modules are dispatched on the CPU or the disk. Make sure you have enough GPU RAM to fit the quantized model. If you want to dispatch the model on the CPU or the disk while keeping these modul…
qwen3_moemultitest_model_15b_a2b_long_prompt_sdpaload_error7/7(line 74) ValueError: Some modules are dispatched on the CPU or the disk. Make sure you have enough GPU RAM to fit the quantized model. If you want to dispatch the model on the CPU or the disk while keeping these modul…
qwen3_moemultitest_speculative_generationload_error7/7(line 74) ValueError: Some modules are dispatched on the CPU or the disk. Make sure you have enough GPU RAM to fit the quantized model. If you want to dispatch the model on the CPU or the disk while keeping these modul…
bridgetowermultitest_constrastive_learningother7/7(line 716) ValueError: Unrecognized image processor in BridgeTower/bridgetower-large-itm-mlm-itc. Should have a `image_processor_type` key in its preprocessor_config.json of config.json, or one of the following `model_…
bridgetowermultitest_image_and_text_retrievalother7/7(line 716) ValueError: Unrecognized image processor in BridgeTower/bridgetower-base-itm-mlm. Should have a `image_processor_type` key in its preprocessor_config.json of config.json, or one of the following `model_type`…
bridgetowermultitest_masked_language_modelingother7/7(line 716) ValueError: Unrecognized image processor in BridgeTower/bridgetower-base-itm-mlm. Should have a `image_processor_type` key in its preprocessor_config.json of config.json, or one of the following `model_type`…
bridgetowersingletest_constrastive_learningother7/7(line 716) ValueError: Unrecognized image processor in BridgeTower/bridgetower-large-itm-mlm-itc. Should have a `image_processor_type` key in its preprocessor_config.json of config.json, or one of the following `model_…
bridgetowersingletest_image_and_text_retrievalother7/7(line 716) ValueError: Unrecognized image processor in BridgeTower/bridgetower-base-itm-mlm. Should have a `image_processor_type` key in its preprocessor_config.json of config.json, or one of the following `model_type`…
bridgetowersingletest_masked_language_modelingother7/7(line 716) ValueError: Unrecognized image processor in BridgeTower/bridgetower-base-itm-mlm. Should have a `image_processor_type` key in its preprocessor_config.json of config.json, or one of the following `model_type`…
cohere2_visionmultitest_model_forward_visionother7/7(line 595) OSError: Repo id must be in the form 'repo_name' or 'namespace/repo_name': '/root/repos/moe/engines/command_a+_bf16'. Use `repo_type` argument if needed.
cohere2_visionmultitest_model_generate_visionother7/7(line 595) OSError: Repo id must be in the form 'repo_name' or 'namespace/repo_name': '/root/repos/moe/engines/command_a+_bf16'. Use `repo_type` argument if needed.
cohere2_visionsingletest_model_forward_visionother7/7(line 595) OSError: Repo id must be in the form 'repo_name' or 'namespace/repo_name': '/root/repos/moe/engines/command_a+_bf16'. Use `repo_type` argument if needed.
cohere2_visionsingletest_model_generate_visionother7/7(line 595) OSError: Repo id must be in the form 'repo_name' or 'namespace/repo_name': '/root/repos/moe/engines/command_a+_bf16'. Use `repo_type` argument if needed.
colqwen2multitest_model_integration_testother7/7(line 114) TypeError: only a single or a list of entries is supported but got type=<class 'datasets.arrow_dataset.Column'>
colqwen2singletest_model_integration_testother7/7(line 114) TypeError: only a single or a list of entries is supported but got type=<class 'datasets.arrow_dataset.Column'>
deepseek_vl_hybridmultitest_model_text_generationother7/7(line 67) RuntimeError: Expected all tensors to be on the same device, but found at least two devices, cuda:0 and cuda:1!
deepseek_vl_hybridsingletest_model_text_generation_with_multi_imageother7/7(line 470) RuntimeError: You can't move a model that has some modules offloaded to cpu or disk.
generationmultitest_cache_device_map_with_vision_layer_device_mapother7/7(line 1644) ValueError: The device_map provided does not give any device for the following parameters: model.vision_tower.embeddings.patch_embedding.weight, model.vision_tower.embeddings.patch_embedding.bias, model.vis…
generationmultitest_generate_multi_accelerator_causal_maskother7/7(line 1644) ValueError: The device_map provided does not give any device for the following parameters: model.visual.patch_embed.proj.weight, model.visual.blocks.0.norm1.weight, model.visual.blocks.0.norm1.bias, model.v…
generationsingletest_cache_device_map_with_vision_layer_device_mapother7/7(line 1644) ValueError: The device_map provided does not give any device for the following parameters: model.vision_tower.embeddings.patch_embedding.weight, model.vision_tower.embeddings.patch_embedding.bias, model.vis…
glm_imagemultitest_image_to_image_generationother7/7(line 322) RuntimeError: The size of tensor a (864) must match the size of tensor b (630) at non-singleton dimension 0
glm_imagesingletest_image_to_image_generationother7/7(line 322) RuntimeError: The size of tensor a (864) must match the size of tensor b (630) at non-singleton dimension 0
janusmultitest_model_generate_imagesother7/7(line 1247) TypeError: GenerationMixin._prepare_static_cache() missing 1 required positional argument: 'prefill_chunk_size'
janusmultitest_model_text_generationother7/7(line 2118) ValueError: Image features and image tokens do not match, tokens: 0, features: 1179648
janusmultitest_model_text_generation_with_multi_imageother7/7(line 2118) ValueError: Image features and image tokens do not match, tokens: 0, features: 2359296
janussingletest_model_generate_imagesother7/7(line 1247) TypeError: GenerationMixin._prepare_static_cache() missing 1 required positional argument: 'prefill_chunk_size'
janussingletest_model_text_generationother7/7(line 2118) ValueError: Image features and image tokens do not match, tokens: 0, features: 1179648
janussingletest_model_text_generation_with_multi_imageother7/7(line 2118) ValueError: Image features and image tokens do not match, tokens: 0, features: 2359296
kimi_k25multitest_model_logitsother7/7(line 552) RuntimeError: indices should be either on cpu or on the same device as the indexed tensor (cuda:1)
kimi_k25multitest_model_logits_batchedother7/7(line 552) RuntimeError: indices should be either on cpu or on the same device as the indexed tensor (cuda:1)
minicpm3multitest_minicpm3_4b_generationother7/7(line 199) KeyError: 'factor'
minicpm3multitest_minicpm3_4b_logitsother7/7(line 199) KeyError: 'factor'
minicpm3singletest_minicpm3_4b_generationother7/7(line 199) KeyError: 'factor'
minicpm3singletest_minicpm3_4b_logitsother7/7(line 199) KeyError: 'factor'
minicpmv4_6multitest_small_model_video_generationother7/7(line 384) RuntimeError: shape '[49, 1014, 1152]' is invalid for input of size 57286656
minicpmv4_6singletest_small_model_video_generationother7/7(line 384) RuntimeError: shape '[49, 1014, 1152]' is invalid for input of size 57286656
peft_integrationmultitest_hotswap_with_compile_and_higher_rank_worksother7/7(line 316) RuntimeError: You set `ignore_mismatched_sizes` to `False`, thus raising an error. For details look at the above report!
peft_integrationmultitest_hotswap_with_compile_and_lower_rank_worksother7/7(line 316) RuntimeError: You set `ignore_mismatched_sizes` to `False`, thus raising an error. For details look at the above report!
peft_integrationmultitest_hotswap_without_compile_and_with_higher_rank_worksother7/7(line 316) RuntimeError: You set `ignore_mismatched_sizes` to `False`, thus raising an error. For details look at the above report!
peft_integrationmultitest_hotswap_without_compile_and_with_lower_rank_worksother7/7(line 316) RuntimeError: You set `ignore_mismatched_sizes` to `False`, thus raising an error. For details look at the above report!
peft_integrationsingletest_hotswap_with_compile_and_higher_rank_worksother7/7(line 316) RuntimeError: You set `ignore_mismatched_sizes` to `False`, thus raising an error. For details look at the above report!
peft_integrationsingletest_hotswap_with_compile_and_lower_rank_worksother7/7(line 316) RuntimeError: You set `ignore_mismatched_sizes` to `False`, thus raising an error. For details look at the above report!
peft_integrationsingletest_hotswap_without_compile_and_with_higher_rank_worksother7/7(line 316) RuntimeError: You set `ignore_mismatched_sizes` to `False`, thus raising an error. For details look at the above report!
peft_integrationsingletest_hotswap_without_compile_and_with_lower_rank_worksother7/7(line 316) RuntimeError: You set `ignore_mismatched_sizes` to `False`, thus raising an error. For details look at the above report!
pegasusmultitest_pegasus_xsum_summaryother7/7(line 350) assert torch.Size([2, 422]) == (2, 421)
pegasussingletest_pegasus_xsum_summaryother7/7(line 350) assert torch.Size([2, 422]) == (2, 421)
phi3multitest_export_static_cacheother7/7(line 1507) torch._dynamo.exc.Unsupported: Data-dependent branching
phi3singletest_export_static_cacheother7/7(line 1507) torch._dynamo.exc.Unsupported: Data-dependent branching
qwen2_moemultitest_speculative_generationother7/7(line 311) RuntimeError: We encountered some issues during automatic conversion of the weights. For details look at the `CONVERSION` entries of the above report!
qwen2_moesingletest_speculative_generationother7/7(line 311) RuntimeError: We encountered some issues during automatic conversion of the weights. For details look at the `CONVERSION` entries of the above report!
qwen3_vl_moemultitest_small_model_integration_testother7/7(line 311) RuntimeError: We encountered some issues during automatic conversion of the weights. For details look at the `CONVERSION` entries of the above report!
qwen3_vl_moemultitest_small_model_integration_test_batchother7/7(line 311) RuntimeError: We encountered some issues during automatic conversion of the weights. For details look at the `CONVERSION` entries of the above report!
qwen3_vl_moemultitest_small_model_integration_test_batch_different_resolutionsother7/7(line 311) RuntimeError: We encountered some issues during automatic conversion of the weights. For details look at the `CONVERSION` entries of the above report!
qwen3_vl_moemultitest_small_model_integration_test_batch_wo_imageother7/7(line 311) RuntimeError: We encountered some issues during automatic conversion of the weights. For details look at the `CONVERSION` entries of the above report!
qwen3_vl_moemultitest_small_model_integration_test_expandother7/7(line 311) RuntimeError: We encountered some issues during automatic conversion of the weights. For details look at the `CONVERSION` entries of the above report!
qwen3_vl_moemultitest_small_model_integration_test_expand_with_videoother7/7(line 311) RuntimeError: We encountered some issues during automatic conversion of the weights. For details look at the `CONVERSION` entries of the above report!
recurrent_gemmamultitest_model_2b_8bitother7/7(line 158) RuntimeError: (*bias): last dimension must be contiguous
recurrent_gemmasingletest_model_2b_8bitother7/7(line 158) RuntimeError: (*bias): last dimension must be contiguous
seamless_m4tmultitest_speech_to_speech_modelother7/7(line 421) ValueError: Invalid input type. Must be a single audio or a list of audio
seamless_m4tmultitest_speech_to_text_modelother7/7(line 421) ValueError: Invalid input type. Must be a single audio or a list of audio
seamless_m4tmultitest_to_rus_speechother7/7(line 421) ValueError: Invalid input type. Must be a single audio or a list of audio
seamless_m4tsingletest_speech_to_speech_modelother7/7(line 421) ValueError: Invalid input type. Must be a single audio or a list of audio
seamless_m4tsingletest_speech_to_text_modelother7/7(line 421) ValueError: Invalid input type. Must be a single audio or a list of audio
seamless_m4tsingletest_to_rus_speechother7/7(line 421) ValueError: Invalid input type. Must be a single audio or a list of audio
utilsmultitest_chunked_prefill_initializes_static_cache_eagerlyother7/7(line 331) TypeError: Pointer argument must be either uint64 or have data_ptr method
utilsmultitest_data_parallel_dynamic_cacheother7/7(line 1834) RuntimeError: a Tensor with 2 elements cannot be converted to Scalar
utilssingletest_chunked_prefill_initializes_static_cache_eagerlyother7/7(line 331) TypeError: Pointer argument must be either uint64 or have data_ptr method
bambamultitest_simple_batched_generate_with_paddingoutput_mismatch7/7(line 627) AssertionError: '<|be[35 chars]on this lovely evening? I hope you are all doing well. Today I' != '<|be[35 chars]on this lovely evening? I hope you are doing well. I am here'
bambamultitest_simple_generateoutput_mismatch7/7(line 575) AssertionError: '<|be[35 chars]on this lovely evening? I hope you are all doing well. Today I' != '<|be[35 chars]on this lovely evening? I hope you are all having a good time.'
bambasingletest_simple_batched_generate_with_paddingoutput_mismatch7/7(line 627) AssertionError: '<|be[35 chars]on this lovely evening? I hope you are all doing well. Today I' != '<|be[35 chars]on this lovely evening? I hope you are doing well. I am here'
bambasingletest_simple_generateoutput_mismatch7/7(line 575) AssertionError: '<|be[35 chars]on this lovely evening? I hope you are all doing well. Today I' != '<|be[35 chars]on this lovely evening? I hope you are all having a good time.'
big_birdmultitest_fill_maskoutput_mismatch7/7(line 921) AssertionError: '' != 'happiness'
big_birdsingletest_fill_maskoutput_mismatch7/7(line 921) AssertionError: '' != 'happiness'
bitnetmultitest_model_logitsoutput_mismatch7/7(line 198) AssertionError: Tensor-likes are not close!
bitnetsingletest_model_logitsoutput_mismatch7/7(line 198) AssertionError: Tensor-likes are not close!
blip_2multitest_inference_t5output_mismatch7/7(line 1658) AssertionError: Lists differ: [0, 2335, 1556, 28, 1782, 30, 8, 2608, 1] != [0, 3, 9, 2335, 19, 1556, 28, 160, 1782, 30, 8, 2608, 1]
blip_2multitest_inference_t5_batched_beam_searchoutput_mismatch7/7(line 1713) AssertionError: Lists differ: [0, 2335, 1556, 28, 1782, 30, 8, 2608, 1] != [0, 3, 9, 2335, 19, 1556, 28, 160, 1782, 30, 8, 2608, 1]
blip_2multitest_inference_t5_multi_acceleratoroutput_mismatch7/7(line 1782) AssertionError: Lists differ: [0, 2335, 1556, 28, 1782, 30, 8, 2608, 1] != [0, 3, 9, 2335, 19, 1556, 28, 160, 1782, 30, 8, 2608, 1]
blip_2singletest_inference_t5output_mismatch7/7(line 1658) AssertionError: Lists differ: [0, 2335, 1556, 28, 1782, 30, 8, 2608, 1] != [0, 3, 9, 2335, 19, 1556, 28, 160, 1782, 30, 8, 2608, 1]
blip_2singletest_inference_t5_batched_beam_searchoutput_mismatch7/7(line 1713) AssertionError: Lists differ: [0, 2335, 1556, 28, 1782, 30, 8, 2608, 1] != [0, 3, 9, 2335, 19, 1556, 28, 160, 1782, 30, 8, 2608, 1]
bloommultitest_batch_generated_textoutput_mismatch7/7(line 621) AssertionError: Lists differ: ['Hello what is', 'Running a quick test with the followi[54 chars]the'] != ['Hello what is the best way to get the data from the se[127 chars]on2']
bloommultitest_batch_generation_paddingoutput_mismatch7/7(line 586) AssertionError: Lists differ: [5941[15 chars]632, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0,[82 chars]0, 0] != [5941[15 chars]632, 419, 682, 15, 473, 912, 267, 40704, 15, 1[186 chars] 912]
bloommultitest_simple_generationoutput_mismatch7/7(line 539) AssertionError: 'I en[58 chars] play. I am a very active person, and I am a v[75 chars]am a' != 'I en[58 chars] play with the kids. I am a very active person[86 chars]nd I'
bloomsingletest_batch_generated_textoutput_mismatch7/7(line 621) AssertionError: Lists differ: ['Hello what is', 'Running a quick test with the followi[54 chars]the'] != ['Hello what is the best way to get the data from the se[127 chars]on2']
bloomsingletest_batch_generation_paddingoutput_mismatch7/7(line 586) AssertionError: Lists differ: [5941[15 chars]632, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0, 0,[82 chars]0, 0] != [5941[15 chars]632, 419, 682, 15, 473, 912, 267, 40704, 15, 1[186 chars] 912]
bloomsingletest_simple_generationoutput_mismatch7/7(line 539) AssertionError: 'I en[58 chars] play. I am a very active person, and I am a v[75 chars]am a' != 'I en[58 chars] play with the kids. I am a very active person[86 chars]nd I'
cohere2multitest_model_pipeline_bf16output_mismatch7/7(line 160) AssertionError: 'Hell[22 chars] for my school and I need to create a website [31 chars] the' != 'Hell[22 chars] for a school assignment and I need to create [37 chars]have'
cohere2singletest_model_pipeline_bf16output_mismatch7/7(line 160) AssertionError: 'Hell[22 chars] for my school and I need to create a website [31 chars] the' != 'Hell[22 chars] for a school assignment and I need to create [37 chars]have'
cohere2_visionsingletest_model_integration_forwardoutput_mismatch7/7(line 687) AssertionError: False is not true : Actual logits: tensor([2.3711, 1.6689, 1.8389, 1.9785, 1.9131], dtype=torch.float16)
colqwen2multitest_model_integration_test_2output_mismatch7/7(line 405) AssertionError: Expected scores tensor([[16.3750, 10.9375, 14.7500],
colqwen2singletest_model_integration_test_2output_mismatch7/7(line 405) AssertionError: Expected scores tensor([[16.3750, 10.9375, 14.7500],
convnextv2multitest_inference_image_classification_headoutput_mismatch7/7(line 315) AssertionError: Tensor-likes are not close!
convnextv2singletest_inference_image_classification_headoutput_mismatch7/7(line 315) AssertionError: Tensor-likes are not close!
cvtmultitest_inference_image_classification_headoutput_mismatch7/7(line 278) AssertionError: Tensor-likes are not close!
cvtsingletest_inference_image_classification_headoutput_mismatch7/7(line 278) AssertionError: Tensor-likes are not close!
dab_detrmultitest_inference_object_detection_headoutput_mismatch7/7(line 800) AssertionError: Tensor-likes are not close!
dab_detrsingletest_inference_object_detection_headoutput_mismatch7/7(line 800) AssertionError: Tensor-likes are not close!
deepseek_vlmultitest_model_text_generation_batchedoutput_mismatch7/7(line 149) AssertionError: Lists differ: ['You[222 chars]tant:The image depicts a snowy landscape with [367 chars]the'] != ['You[222 chars]tant:What is a cat, a cat, a cat, a cat, a cat[329 chars]the']
deepseek_vlsingletest_model_text_generation_batchedoutput_mismatch7/7(line 149) AssertionError: Lists differ: ['You[222 chars]tant:The image depicts a snowy landscape with [367 chars]the'] != ['You[222 chars]tant:What is a cat, a cat, a cat, a cat, a cat[329 chars]the']
deepseek_vl_hybridsingletest_model_text_generation_batchedoutput_mismatch7/7(line 377) AssertionError: Lists differ: ['You[224 chars]nt:\nThe image depicts a fluffy, light brown a[371 chars]he '] != ['You[224 chars]nt:\nA fluffy animal in a fluffyThe image,The [329 chars]he ']
diamultitest_dia_model_integration_generate_audio_contextoutput_mismatch7/7(line 738) AssertionError: Tensor-likes are not equal!
diasingletest_dia_model_integration_generate_audio_contextoutput_mismatch7/7(line 738) AssertionError: Tensor-likes are not equal!
diffllamamultitest_compile_static_cacheoutput_mismatch7/7(line 484) AssertionError: Lists differ: ['Sim[41 chars]that 1) the speed of light is constant in all [301 chars]y p'] != ['Sim[41 chars]that 2.5 5 5 5 5 5 5 5 5 5 5 5 5 5 5 5 5 5 5 '[133 chars]a a']
diffllamasingletest_compile_static_cacheoutput_mismatch7/7(line 484) AssertionError: Lists differ: ['Sim[41 chars]that 1) the speed of light is constant in all [301 chars]y p'] != ['Sim[41 chars]that 2.5 5 5 5 5 5 5 5 5 5 5 5 5 5 5 5 5 5 5 '[133 chars]a a']
diffusion_gemmamultitest_diffusion_gemma_chat_templateoutput_mismatch7/7(line 897) AssertionError: Lists differ: [[2, [46 chars]23613, 236761, 106, 107, 105, 4368, 107]] != [[2, [46 chars]23613, 236761, 106, 107, 105, 4368, 107, 100, 45518, 107, 101]]
diffusion_gemmamultitest_minified_diffusion_gemma_forward_batchedoutput_mismatch7/7(line 1694) AssertionError: Tensor-likes are not close!
diffusion_gemmamultitest_minified_diffusion_gemma_forward_text_onlyoutput_mismatch7/7(line 1507) AssertionError: Tensor-likes are not close!
diffusion_gemmamultitest_minified_diffusion_gemma_forward_with_imageoutput_mismatch7/7(line 1600) AssertionError: Tensor-likes are not close!
diffusion_gemmasingletest_diffusion_gemma_chat_templateoutput_mismatch7/7(line 897) AssertionError: Lists differ: [[2, [46 chars]23613, 236761, 106, 107, 105, 4368, 107]] != [[2, [46 chars]23613, 236761, 106, 107, 105, 4368, 107, 100, 45518, 107, 101]]
diffusion_gemmasingletest_minified_diffusion_gemma_forward_batchedoutput_mismatch7/7(line 1694) AssertionError: Tensor-likes are not close!
diffusion_gemmasingletest_minified_diffusion_gemma_forward_text_onlyoutput_mismatch7/7(line 1507) AssertionError: Tensor-likes are not close!
diffusion_gemmasingletest_minified_diffusion_gemma_forward_with_imageoutput_mismatch7/7(line 1600) AssertionError: Tensor-likes are not close!
efficientnetmultitest_inference_image_classification_headoutput_mismatch7/7(line 259) AssertionError: Tensor-likes are not close!
efficientnetsingletest_inference_image_classification_headoutput_mismatch7/7(line 259) AssertionError: Tensor-likes are not close!
eomt_dinov3multitest_inference_bf16output_mismatch7/7(line 315) AssertionError: Tensor-likes are not close!
eomt_dinov3singletest_inference_bf16output_mismatch7/7(line 315) AssertionError: Tensor-likes are not close!
exaone4multitest_model_logitsoutput_mismatch7/7(line 88) AssertionError: Tensor-likes are not close!
exaone4singletest_model_generation_beyond_sliding_windowoutput_mismatch7/7(line 149) AssertionError: " Thi[46 chars] and the atmosphere is so relaxing. I'm gratef[47 chars]. It" != " Thi[46 chars] and I'm grateful for the opportunity to exper[26 chars]reak"
exaone4singletest_model_logitsoutput_mismatch7/7(line 88) AssertionError: Tensor-likes are not close!
exaone_moemultitest_model_logitsoutput_mismatch7/7(line 110) AssertionError: Tensor-likes are not close!
exaone_moesingletest_model_logitsoutput_mismatch7/7(line 110) AssertionError: Tensor-likes are not close!
fastspeech2_conformermultitest_training_integrationoutput_mismatch7/7(line 453) AssertionError: Tensor-likes are not close!
fastspeech2_conformersingletest_training_integrationoutput_mismatch7/7(line 453) AssertionError: Tensor-likes are not close!
fuyumultitest_greedy_generationoutput_mismatch7/7(line 295) AssertionError: '\x04 A bus parked on the side of a road.' != 'A blue bus parked on the side of a road.'
fuyusingletest_greedy_generationoutput_mismatch7/7(line 295) AssertionError: '\x04 A bus parked on the side of a road.' != 'A blue bus parked on the side of a road.'
generationmultitest_beam_search_early_stop_heuristicoutput_mismatch7/7(line 3421) AssertionError: '<|user|>\nWhat is 3+5?\n<|assistant|>\nT[215 chars]d 8.' != "<|user|>\nWhat is 3+5?\n<|assistant|>\nT[282 chars]\\)."
generationmultitest_mtp_use_correct_device_when_draftingoutput_mismatch7/7(line 4095) AssertionError: device(type='cuda', index=0) != device(type='cuda', index=1)
generationsingletest_beam_search_early_stop_heuristicoutput_mismatch7/7(line 3421) AssertionError: '<|user|>\nWhat is 3+5?\n<|assistant|>\nT[215 chars]d 8.' != "<|user|>\nWhat is 3+5?\n<|assistant|>\nT[282 chars]\\)."
granitemultitest_model_3b_logits_bf16output_mismatch7/7(line 687) AssertionError: False is not true
granitesingletest_model_3b_logits_bf16output_mismatch7/7(line 687) AssertionError: False is not true
grounding_dinomultitest_cross_attention_maskoutput_mismatch7/7(line 787) AssertionError: Tensor-likes are not close!
grounding_dinomultitest_grounding_dino_lossoutput_mismatch7/7(line 869) AssertionError: Scalars are not close!
grounding_dinomultitest_inference_object_detection_headoutput_mismatch7/7(line 678) AssertionError: Tensor-likes are not close!
grounding_dinomultitest_inference_object_detection_head_equivalence_cpu_acceleratoroutput_mismatch7/7(line 745) AssertionError: Tensor-likes are not close!
grounding_dinosingletest_cross_attention_maskoutput_mismatch7/7(line 787) AssertionError: Tensor-likes are not close!
grounding_dinosingletest_grounding_dino_lossoutput_mismatch7/7(line 869) AssertionError: Scalars are not close!
grounding_dinosingletest_inference_object_detection_headoutput_mismatch7/7(line 678) AssertionError: Tensor-likes are not close!
grounding_dinosingletest_inference_object_detection_head_equivalence_cpu_acceleratoroutput_mismatch7/7(line 745) AssertionError: Tensor-likes are not close!
hrm_textmultitest_forward_logitsoutput_mismatch7/7(line 280) AssertionError: Tensor-likes are not close!
hrm_textsingletest_forward_logitsoutput_mismatch7/7(line 280) AssertionError: Tensor-likes are not close!
hy_v3multitest_small_model_logits_batchedoutput_mismatch7/7(line 105) AssertionError: Tensor-likes are not close!
hy_v3singletest_small_model_logits_batchedoutput_mismatch7/7(line 105) AssertionError: Tensor-likes are not close!
jais2multitest_model_logitsoutput_mismatch7/7(line 118) AssertionError: Tensor-likes are not close!
jais2singletest_model_logitsoutput_mismatch7/7(line 118) AssertionError: Tensor-likes are not close!
jambamultitest_simple_batched_generate_with_paddingoutput_mismatch7/7(line 582) AssertionError: "<|startoftext|>Tell me a story<|pad|><|p[50 chars]t I'" != '<|pad|><|pad|><|pad|><|pad|><|pad|><|pad[76 chars]ates'
jambasingletest_simple_batched_generate_with_paddingoutput_mismatch7/7(line 582) AssertionError: "<|startoftext|>Tell me a story<|pad|><|p[50 chars]t I'" != '<|pad|><|pad|><|pad|><|pad|><|pad|><|pad[76 chars]ates'
kimi_k25singletest_model_logitsoutput_mismatch7/7(line 324) AssertionError: Tensor-likes are not close!
kimi_k25singletest_model_logits_batchedoutput_mismatch7/7(line 389) AssertionError: Tensor-likes are not close!
kosmos2multitest_inference_interpolate_pos_encodingoutput_mismatch7/7(line 783) AssertionError: Tensor-likes are not close!
kosmos2multitest_snowman_image_captioningoutput_mismatch7/7(line 550) AssertionError:
kosmos2multitest_snowman_image_captioning_batchoutput_mismatch7/7(line 712) AssertionError: Lists differ: ['<gr[35 chars]ail: A snowman is sitting in front of a fire, [575 chars]t>.'] != ['<gr[35 chars]ail: The image features a snowman sitting by<p[836 chars]t>.']
kosmos2singletest_inference_interpolate_pos_encodingoutput_mismatch7/7(line 783) AssertionError: Tensor-likes are not close!
kosmos2singletest_snowman_image_captioningoutput_mismatch7/7(line 550) AssertionError:
kosmos2singletest_snowman_image_captioning_batchoutput_mismatch7/7(line 712) AssertionError: Lists differ: ['<gr[35 chars]ail: A snowman is sitting in front of a fire, [575 chars]t>.'] != ['<gr[35 chars]ail: The image features a snowman sitting by<p[836 chars]t>.']
kosmos2_5multitest_eageroutput_mismatch7/7(line 585) AssertionError: Lists differ: ['<bb[65 chars]<y_612></bbox>[REG] BLACK SAKURA\n<bbox><x_690[603 chars]0\n'] != ['<bb[65 chars]<y_611></bbox>[REG] BLACK SAKURA\n<bbox><x_690[603 chars]0\n']
kosmos2_5singletest_eageroutput_mismatch7/7(line 585) AssertionError: Lists differ: ['<bb[65 chars]<y_612></bbox>[REG] BLACK SAKURA\n<bbox><x_690[603 chars]0\n'] != ['<bb[65 chars]<y_611></bbox>[REG] BLACK SAKURA\n<bbox><x_690[603 chars]0\n']
layoutlmv2multitest_processor_case_1output_mismatch7/7(line 213) AssertionError: Sequences differ: "[CLS[522 chars]t itc ' s new fmcg businesses are the fastest [829 chars]PAD]" != "[CLS[522 chars]t itc's new fmcg businesses are the fastest gr[827 chars]PAD]"
layoutlmv2multitest_processor_case_4output_mismatch7/7(line 358) AssertionError: Sequences differ: "[CLS] what ' s his name? [SEP] 11 : 14 to 11 : 39 a[1108 chars]SEP]" != "[CLS] what's his name? [SEP] 11 : 14 to 11 : 39 a. [1106 chars]SEP]"
layoutlmv2multitest_processor_case_5output_mismatch7/7(line 406) AssertionError: Sequences differ: "[CLS] what ' s his name? [SEP] hello world [SEP]" != "[CLS] what's his name? [SEP] hello world [SEP]"
layoutlmv2singletest_processor_case_1output_mismatch7/7(line 213) AssertionError: Sequences differ: "[CLS[522 chars]t itc ' s new fmcg businesses are the fastest [829 chars]PAD]" != "[CLS[522 chars]t itc's new fmcg businesses are the fastest gr[827 chars]PAD]"
layoutlmv2singletest_processor_case_4output_mismatch7/7(line 358) AssertionError: Sequences differ: "[CLS] what ' s his name? [SEP] 11 : 14 to 11 : 39 a[1108 chars]SEP]" != "[CLS] what's his name? [SEP] 11 : 14 to 11 : 39 a. [1106 chars]SEP]"
layoutlmv2singletest_processor_case_5output_mismatch7/7(line 406) AssertionError: Sequences differ: "[CLS] what ' s his name? [SEP] hello world [SEP]" != "[CLS] what's his name? [SEP] hello world [SEP]"
llava_next_videomultitest_small_model_integration_testoutput_mismatch7/7(line 387) AssertionError: 'USER[154 chars]hile wearing a pair of glasses that are too la[24 chars] are' != 'USER[154 chars]hile another child is attempting to read the s[45 chars]eems'
llava_next_videosingletest_small_model_integration_testoutput_mismatch7/7(line 387) AssertionError: 'USER[154 chars]hile wearing a pair of glasses that are too la[24 chars] are' != 'USER[154 chars]hile another child is attempting to read the s[45 chars]eems'
lukemultitest_inference_base_modeloutput_mismatch7/7(line 905) AssertionError: Tensor-likes are not close!
lukemultitest_inference_large_modeloutput_mismatch7/7(line 940) AssertionError: Tensor-likes are not close!
lukesingletest_inference_base_modeloutput_mismatch7/7(line 905) AssertionError: Tensor-likes are not close!
lukesingletest_inference_large_modeloutput_mismatch7/7(line 940) AssertionError: Tensor-likes are not close!
lw_detrmultitest_inference_object_detection_head_tinyoutput_mismatch7/7(line 690) AssertionError: Tensor-likes are not close!
lw_detrmultitest_inference_object_detection_head_xlargeoutput_mismatch7/7(line 766) AssertionError: Tensor-likes are not close!
lw_detrsingletest_inference_object_detection_head_tinyoutput_mismatch7/7(line 690) AssertionError: Tensor-likes are not close!
lw_detrsingletest_inference_object_detection_head_xlargeoutput_mismatch7/7(line 766) AssertionError: Tensor-likes are not close!
m2m_100multitest_seq_to_seq_generationoutput_mismatch7/7(line 397) AssertionError: assert ['</s>__en__T... France.</s>'] == ['</s> __en__... France.</s>']
m2m_100singletest_seq_to_seq_generationoutput_mismatch7/7(line 397) AssertionError: assert ['</s>__en__T... France.</s>'] == ['</s> __en__... France.</s>']
mimimultitest_integrationoutput_mismatch7/7(line 687) AssertionError: np.False_ is not true
mimimultitest_integration_longformoutput_mismatch7/7(line 687) AssertionError: np.False_ is not true
mimisingletest_integrationoutput_mismatch7/7(line 687) AssertionError: np.False_ is not true
mimisingletest_integration_longformoutput_mismatch7/7(line 687) AssertionError: np.False_ is not true
ministralmultitest_model_8b_logitsoutput_mismatch7/7(line 93) AssertionError: Tensor-likes are not close!
ministralsingletest_model_8b_logitsoutput_mismatch7/7(line 93) AssertionError: Tensor-likes are not close!
ministral3multitest_model_3b_generationoutput_mismatch7/7(line 130) AssertionError: 'My favourite condiment is 100% pure oliv[64 chars] can' != "My favourite condiment is 100% pure oliv[46 chars]t in"
ministral3multitest_model_3b_logitsoutput_mismatch7/7(line 102) AssertionError: Tensor-likes are not close!
ministral3singletest_model_3b_generationoutput_mismatch7/7(line 130) AssertionError: 'My favourite condiment is 100% pure oliv[64 chars] can' != "My favourite condiment is 100% pure oliv[46 chars]t in"
ministral3singletest_model_3b_logitsoutput_mismatch7/7(line 102) AssertionError: Tensor-likes are not close!
mixtralmultitest_small_model_logitsoutput_mismatch7/7(line 143) AssertionError: Tensor-likes are not close!
mixtralmultitest_small_model_logits_batchedoutput_mismatch7/7(line 188) AssertionError: Tensor-likes are not close!
mixtralsingletest_small_model_logitsoutput_mismatch7/7(line 143) AssertionError: Tensor-likes are not close!
mixtralsingletest_small_model_logits_batchedoutput_mismatch7/7(line 188) AssertionError: Tensor-likes are not close!
mllamamultitest_11b_model_integration_batched_generateoutput_mismatch7/7(line 699) AssertionError: "This image showst's a photo of a person named I'm not abl[59 chars]ge's" != "This image shows\nI'm not able to provide information on [64 chars]ning"
mllamamultitest_11b_model_integration_multi_image_generateoutput_mismatch7/7(line 762) AssertionError: 'The image shows a red octagonal stop sign w[59 chars]to a' != 'This image shows a long wooden dock extendi[67 chars]ling'
mllamasingletest_11b_model_integration_batched_generateoutput_mismatch7/7(line 699) AssertionError: "This image showst's a photo of a person named I'm not abl[59 chars]ge's" != "This image shows\nI'm not able to provide information on [64 chars]ning"
mllamasingletest_11b_model_integration_multi_image_generateoutput_mismatch7/7(line 762) AssertionError: 'The image shows a red octagonal stop sign w[59 chars]to a' != 'This image shows a long wooden dock extendi[67 chars]ling'
mlukemultitest_entity_classification_no_padding_or_truncationoutput_mismatch7/7(line 453) AssertionError: '<s> Japanese is an<s> East Asian language<s> spoken by about[40 chars]</s>' != '<s> Japanese is an<ent>East Asian language<ent>spoken by abo[42 chars]</s>'
mlukemultitest_entity_pair_classification_no_padding_or_truncationoutput_mismatch7/7(line 507) AssertionError: '<s><s> Japanese<s> is an East Asian language [64 chars]</s>' != '<s><ent>Japanese<ent>is an East Asian languag[68 chars]</s>'
mlukemultitest_entity_span_classification_no_padding_or_truncationoutput_mismatch7/7(line 572) AssertionError: '<s> [33 chars]e spoken by about 128 million people, primarily in Japan .</s>' != '<s> [33 chars]e spoken by about 128 million people, primarily in Japan.</s>'
mlukesingletest_entity_classification_no_padding_or_truncationoutput_mismatch7/7(line 453) AssertionError: '<s> Japanese is an<s> East Asian language<s> spoken by about[40 chars]</s>' != '<s> Japanese is an<ent>East Asian language<ent>spoken by abo[42 chars]</s>'
mlukesingletest_entity_pair_classification_no_padding_or_truncationoutput_mismatch7/7(line 507) AssertionError: '<s><s> Japanese<s> is an East Asian language [64 chars]</s>' != '<s><ent>Japanese<ent>is an East Asian languag[68 chars]</s>'
mlukesingletest_entity_span_classification_no_padding_or_truncationoutput_mismatch7/7(line 572) AssertionError: '<s> [33 chars]e spoken by about 128 million people, primarily in Japan .</s>' != '<s> [33 chars]e spoken by about 128 million people, primarily in Japan.</s>'
mm_grounding_dinomultitest_inference_object_detection_headoutput_mismatch7/7(line 672) AssertionError: Tensor-likes are not close!
mm_grounding_dinomultitest_inference_object_detection_head_equivalence_cpu_gpuoutput_mismatch7/7(line 738) AssertionError: Tensor-likes are not close!
mm_grounding_dinomultitest_mm_grounding_dino_lossoutput_mismatch7/7(line 687) AssertionError: False is not true
mm_grounding_dinosingletest_inference_object_detection_headoutput_mismatch7/7(line 672) AssertionError: Tensor-likes are not close!
mm_grounding_dinosingletest_inference_object_detection_head_equivalence_cpu_gpuoutput_mismatch7/7(line 738) AssertionError: Tensor-likes are not close!
mm_grounding_dinosingletest_mm_grounding_dino_lossoutput_mismatch7/7(line 687) AssertionError: False is not true
moshimultitest_moshika_greedy_unconditional_fp16output_mismatch7/7(line 687) AssertionError: np.False_ is not true
moshimultitest_moshiko_greedy_unconditional_fp16output_mismatch7/7(line 687) AssertionError: np.False_ is not true
moshimultitest_moshiko_greedy_unconditional_fp16_eageroutput_mismatch7/7(line 687) AssertionError: False is not true
moshimultitest_moshiko_greedy_unconditional_fp32output_mismatch7/7(line 687) AssertionError: np.False_ is not true
moshisingletest_moshika_greedy_unconditional_fp16output_mismatch7/7(line 687) AssertionError: np.False_ is not true
moshisingletest_moshiko_greedy_unconditional_fp16output_mismatch7/7(line 687) AssertionError: np.False_ is not true
moshisingletest_moshiko_greedy_unconditional_fp16_eageroutput_mismatch7/7(line 687) AssertionError: False is not true
moshisingletest_moshiko_greedy_unconditional_fp32output_mismatch7/7(line 687) AssertionError: np.False_ is not true
musicgenmultitest_generate_text_prompt_samplingoutput_mismatch7/7(line 1263) AssertionError: Tensor-likes are not close!
musicgenmultitest_generate_unconditional_samplingoutput_mismatch7/7(line 1180) AssertionError: Tensor-likes are not close!
musicgensingletest_generate_text_prompt_samplingoutput_mismatch7/7(line 1263) AssertionError: Tensor-likes are not close!
musicgensingletest_generate_unconditional_samplingoutput_mismatch7/7(line 1180) AssertionError: Tensor-likes are not close!
nemotronmultitest_nemotron_8b_generation_eageroutput_mismatch7/7(line 106) AssertionError: Lists differ: ['Wha[46 chars]er: Jupiter\n\nWhat is the answer'] != ['Wha[46 chars]er: Jupiter\n\nWhat is the answer: What is the name of the 19']
nemotronsingletest_nemotron_8b_generation_eageroutput_mismatch7/7(line 106) AssertionError: Lists differ: ['Wha[46 chars]er: Jupiter\n\nWhat is the answer'] != ['Wha[46 chars]er: Jupiter\n\nWhat is the answer: What is the name of the 19']
nllb_moemultitest_inference_logitsoutput_mismatch7/7(line 399) AssertionError: Tensor-likes are not close!
nllb_moesingletest_inference_logitsoutput_mismatch7/7(line 399) AssertionError: Tensor-likes are not close!
oneformermultitest_inference_no_headoutput_mismatch7/7(line 498) AssertionError: Tensor-likes are not close!
oneformermultitest_inference_universal_segmentation_headoutput_mismatch7/7(line 540) AssertionError: Tensor-likes are not close!
oneformersingletest_inference_no_headoutput_mismatch7/7(line 498) AssertionError: Tensor-likes are not close!
oneformersingletest_inference_universal_segmentation_headoutput_mismatch7/7(line 540) AssertionError: Tensor-likes are not close!
ovis2multitest_small_model_integration_test_batch_different_resolutionsoutput_mismatch7/7(line 354) AssertionError: Lists differ: ['sys[81 chars]ant\n', 'system\nYou are a helpful assistant.\[139 chars]et.'] != ['sys[81 chars]ant\nAnswer: I see a brown dog standing on a w[224 chars]et.']
ovis2singletest_small_model_integration_test_batch_different_resolutionsoutput_mismatch7/7(line 354) AssertionError: Lists differ: ['sys[81 chars]ant\n', 'system\nYou are a helpful assistant.\[139 chars]et.'] != ['sys[81 chars]ant\nAnswer: I see a brown dog standing on a w[224 chars]et.']
persimmonmultitest_model_8b_chat_logitsoutput_mismatch7/7(line 99) AssertionError: Tensor-likes are not close!
persimmonsingletest_model_8b_chat_logitsoutput_mismatch7/7(line 99) AssertionError: Tensor-likes are not close!
plbartmultitest_fill_maskoutput_mismatch7/7(line 444) AssertionError: '0 0 the 0 the 0 the 0 the 0 the 0 the 0 the 0 the 0' != '0 0 the 0 the 0 the 0 the 0 the 0 the 0 the 0 the'
plbartmultitest_java_cs_generate_batchoutput_mismatch7/7(line 379) AssertionError: assert ['public int ...turn a * b *'] == ['public int ...rn a * b * c']
plbartmultitest_java_cs_generate_oneoutput_mismatch7/7(line 370) AssertionError: 'public int maximum(int a, int b, int c){return Math.Max(' != 'public int maximum(int a, int b, int c){return Math.Max(a'
plbartsingletest_fill_maskoutput_mismatch7/7(line 444) AssertionError: '0 0 the 0 the 0 the 0 the 0 the 0 the 0 the 0 the 0' != '0 0 the 0 the 0 the 0 the 0 the 0 the 0 the 0 the'
plbartsingletest_java_cs_generate_batchoutput_mismatch7/7(line 379) AssertionError: assert ['public int ...turn a * b *'] == ['public int ...rn a * b * c']
plbartsingletest_java_cs_generate_oneoutput_mismatch7/7(line 370) AssertionError: 'public int maximum(int a, int b, int c){return Math.Max(' != 'public int maximum(int a, int b, int c){return Math.Max(a'
pvtmultitest_inference_image_classificationoutput_mismatch7/7(line 257) AssertionError: Tensor-likes are not close!
pvtmultitest_inference_modeloutput_mismatch7/7(line 284) AssertionError: Tensor-likes are not close!
pvtsingletest_inference_image_classificationoutput_mismatch7/7(line 257) AssertionError: Tensor-likes are not close!
pvtsingletest_inference_modeloutput_mismatch7/7(line 284) AssertionError: Tensor-likes are not close!
pvt_v2multitest_inference_image_classificationoutput_mismatch7/7(line 275) AssertionError: Tensor-likes are not close!
pvt_v2multitest_inference_modeloutput_mismatch7/7(line 292) AssertionError: torch.Size([1, 256, 7, 7]) != torch.Size([1, 50, 512])
pvt_v2singletest_inference_image_classificationoutput_mismatch7/7(line 275) AssertionError: Tensor-likes are not close!
pvt_v2singletest_inference_modeloutput_mismatch7/7(line 292) AssertionError: torch.Size([1, 256, 7, 7]) != torch.Size([1, 50, 512])
qwen2_moemultitest_model_a2_7b_logitsoutput_mismatch7/7(line 157) AssertionError: Tensor-likes are not close!
qwen2_moesingletest_model_a2_7b_logitsoutput_mismatch7/7(line 157) AssertionError: Tensor-likes are not close!
qwen3multitest_model_600m_logitsoutput_mismatch7/7(line 92) AssertionError: Tensor-likes are not close!
qwen3multitest_speculative_generationoutput_mismatch7/7(line 198) AssertionError: 'My f[22 chars]100% beef, 100% beef, 100% beef.' != 'My f[22 chars]100% vegetable oil. It has a rich, creamy text[19 chars]utty'
qwen3singletest_model_600m_logitsoutput_mismatch7/7(line 92) AssertionError: Tensor-likes are not close!
qwen3singletest_speculative_generationoutput_mismatch7/7(line 198) AssertionError: 'My f[22 chars]100% beef, 100% beef, 100% beef.' != 'My f[22 chars]100% vegetable oil. It has a rich, creamy text[19 chars]utty'
qwen3_5multitest_model_video_generationoutput_mismatch7/7(line 852) AssertionError: Lists differ: [248045, 846, 198, 27, 15, 13, 18, 6283, 29, 248053] != [248045, 846, 198, 248053, 27, 15, 13, 18, 6283, 29]
qwen3_5multitest_model_video_generation_batchoutput_mismatch7/7(line 906) AssertionError: Lists differ: [248045, 846, 198, 27, 15, 13, 18, 6283, 29, 248053] != [248045, 846, 198, 248053, 27, 15, 13, 18, 6283, 29]
qwen3_5singletest_model_video_generationoutput_mismatch7/7(line 852) AssertionError: Lists differ: [248045, 846, 198, 27, 15, 13, 18, 6283, 29, 248053] != [248045, 846, 198, 248053, 27, 15, 13, 18, 6283, 29]
qwen3_5singletest_model_video_generation_batchoutput_mismatch7/7(line 906) AssertionError: Lists differ: [248045, 846, 198, 27, 15, 13, 18, 6283, 29, 248053] != [248045, 846, 198, 248053, 27, 15, 13, 18, 6283, 29]
ragsingletest_rag_sequence_generate_batchoutput_mismatch7/7(line 948) AssertionError: Lists differ: [' michael gross', ' monday 17 , 2018', ' te[96 chars]ndo'] != [' albert einstein', ' june 22 , 2018', ' am[85 chars]' 8']
ragsingletest_rag_sequence_generate_batch_from_context_input_idsoutput_mismatch7/7(line 1000) AssertionError: Lists differ: [' michael gross', ' monday 17 , 2018', ' te[96 chars]ndo'] != [' albert einstein', ' june 22 , 2018', ' am[85 chars]' 8']
ragsingletest_rag_sequence_generate_beamoutput_mismatch7/7(line 892) AssertionError: '" in the United States. "People Need Love"[155 chars]hit.' != '"She\'s My Kind of Girl" was released thro[257 chars]nts.'
ragsingletest_rag_token_generate_beamoutput_mismatch7/7(line 854) AssertionError: '"She[14 chars] Girl' != '"She[14 chars] Girl" was released through Epic Records in Ja[179 chars]ses"'
recurrent_gemmamultitest_2b_generateoutput_mismatch7/7(line 254) AssertionError: Lists differ: ['Hel[325 chars]oday the 1990s, the 1990s, the 1990s, the 1990[43 chars]0s,'] != ['Hel[325 chars]oday is a new app that allows you to make mone[256 chars]app']
recurrent_gemmamultitest_2b_sampleoutput_mismatch7/7(line 292) AssertionError: Lists differ: ['Wha[24 chars]Deep Learning (or deep learning) is one of the[107 chars]ple'] != ['Wha[24 chars]Deep learning is the next frontier in computer[98 chars] is']
recurrent_gemmamultitest_longer_than_windowoutput_mismatch7/7(line 340) AssertionError: Lists differ: [' Jean-Philippe Guillet said, "We have no[245 chars]eo.'] != [" Robin's comments follow claims by two m[249 chars]the"]
recurrent_gemmasingletest_2b_generateoutput_mismatch7/7(line 254) AssertionError: Lists differ: ['Hel[325 chars]oday the 1990s, the 1990s, the 1990s, the 1990[43 chars]0s,'] != ['Hel[325 chars]oday is a new app that allows you to make mone[256 chars]app']
recurrent_gemmasingletest_2b_sampleoutput_mismatch7/7(line 292) AssertionError: Lists differ: ['Wha[24 chars]Deep Learning (or deep learning) is one of the[107 chars]ple'] != ['Wha[24 chars]Deep learning is the next frontier in computer[98 chars] is']
recurrent_gemmasingletest_longer_than_windowoutput_mismatch7/7(line 340) AssertionError: Lists differ: [' Jean-Philippe Guillet said, "We have no[245 chars]eo.'] != [" Robin's comments follow claims by two m[249 chars]the"]
reformermultitest_pretrained_generate_crime_and_punishoutput_mismatch7/7(line 1357) AssertionError: 'A fe[36 chars]is ideas, so attentively two or three thousand roubles, and' != 'A fe[36 chars]is ideas, at the first entrance. He was positively for an inst'
reformersingletest_pretrained_generate_crime_and_punishoutput_mismatch7/7(line 1357) AssertionError: 'A fe[36 chars]is ideas, so attentively two or three thousand roubles, and' != 'A fe[36 chars]is ideas, at the first entrance. He was positively for an inst'
regnetmultitest_inference_image_classification_headoutput_mismatch7/7(line 243) AssertionError: Tensor-likes are not close!
regnetsingletest_inference_image_classification_headoutput_mismatch7/7(line 243) AssertionError: Tensor-likes are not close!
seamless_m4t_v2multitest_to_rus_speechoutput_mismatch7/7(line 1061) AssertionError: Lists differ: [3, 256074, 107, 248213, 404, 247792, 247789, 3] != [[3, 256074, 107, 248213, 404, 247792, 247[53 chars], 3]]
seamless_m4t_v2singletest_to_rus_speechoutput_mismatch7/7(line 1061) AssertionError: Lists differ: [3, 256074, 107, 248213, 404, 247792, 247789, 3] != [[3, 256074, 107, 248213, 404, 247792, 247[53 chars], 3]]
smollm3multitest_export_static_cacheoutput_mismatch7/7(line 198) AssertionError: 'Gravity is the force that pulls objects [69 chars] and' != ["Gravity is the force that pulls objects[85 chars] of"]
smollm3multitest_model_3b_logitsoutput_mismatch7/7(line 89) AssertionError: Tensor-likes are not close!
smollm3singletest_export_static_cacheoutput_mismatch7/7(line 198) AssertionError: 'Gravity is the force that pulls objects [69 chars] and' != ["Gravity is the force that pulls objects[85 chars] of"]
smollm3singletest_model_3b_logitsoutput_mismatch7/7(line 89) AssertionError: Tensor-likes are not close!
starcoder2multitest_starcoder2_batched_generation_4bitoutput_mismatch7/7(line 152) AssertionError: Lists differ: ['Hel[207 chars]ld():\n\treturn "Hello World"\n\n@app.route(\'[76 chars]ute'] != ['Hel[207 chars]ld(): hello_world():\n return "Hello World![89 chars]n\n']
starcoder2multitest_starcoder2_batched_generation_eageroutput_mismatch7/7(line 99) AssertionError: Lists differ: ['Hel[223 chars]ld():\n\treturn 'Hello World!'\n\n@app.route('[72 chars]app"] != ['Hel[223 chars]ld(): hello_world():\n return 'Hello World![87 chars]n\n"]
starcoder2multitest_starcoder2_batched_generation_sdpaoutput_mismatch7/7(line 79) AssertionError: Lists differ: ['Hel[223 chars]ld():\n\treturn 'Hello World!'\n\n@app.route('[72 chars]app"] != ['Hel[223 chars]ld(): hello_world():\n return 'Hello World![87 chars]n\n"]
starcoder2singletest_starcoder2_batched_generation_4bitoutput_mismatch7/7(line 152) AssertionError: Lists differ: ['Hel[207 chars]ld():\n\treturn "Hello World"\n\n@app.route(\'[76 chars]ute'] != ['Hel[207 chars]ld(): hello_world():\n return "Hello World![89 chars]n\n']
starcoder2singletest_starcoder2_batched_generation_eageroutput_mismatch7/7(line 99) AssertionError: Lists differ: ['Hel[223 chars]ld():\n\treturn 'Hello World!'\n\n@app.route('[72 chars]app"] != ['Hel[223 chars]ld(): hello_world():\n return 'Hello World![87 chars]n\n"]
starcoder2singletest_starcoder2_batched_generation_sdpaoutput_mismatch7/7(line 79) AssertionError: Lists differ: ['Hel[223 chars]ld():\n\treturn 'Hello World!'\n\n@app.route('[72 chars]app"] != ['Hel[223 chars]ld(): hello_world():\n return 'Hello World![87 chars]n\n"]
superpointmultitest_inferenceoutput_mismatch7/7(line 276) AssertionError: torch.Size([2, 786, 2]) != torch.Size([2, 830, 2])
superpointsingletest_inferenceoutput_mismatch7/7(line 276) AssertionError: torch.Size([2, 786, 2]) != torch.Size([2, 830, 2])
swinv2multitest_inference_fp16output_mismatch7/7(line 487) AssertionError: Tensor-likes are not close!
swinv2singletest_inference_fp16output_mismatch7/7(line 487) AssertionError: Tensor-likes are not close!
t5multitest_compile_static_cacheoutput_mismatch7/7(line 1444) AssertionError: Lists differ: ['the[91 chars]rames . the laws of physics are the same for a[65 chars]t .'] != ['the[91 chars]rames. the laws of physics are the same for al[62 chars]nt.']
t5singletest_compile_static_cacheoutput_mismatch7/7(line 1444) AssertionError: Lists differ: ['the[91 chars]rames . the laws of physics are the same for a[65 chars]t .'] != ['the[91 chars]rames. the laws of physics are the same for al[62 chars]nt.']
univnetmultitest_integrationoutput_mismatch7/7(line 326) AssertionError: Scalars are not close!
univnetsingletest_integrationoutput_mismatch7/7(line 326) AssertionError: Scalars are not close!
utilsmultitest_cache_copyoutput_mismatch7/7(line 670) AssertionError: Lists differ: ['You are a helpful assistant. Help me to [390 chars] is'] != ["You are a helpful assistant. Help me to [385 chars] is']
utilsmultitest_dynamic_cache_hardoutput_mismatch7/7(line 522) AssertionError: "Here[57 chars]ave fur, they have four legs, they have a tail[1045 chars]have" != "Here[57 chars]ave four legs, they have a tail, they have a f[1078 chars]They"
utilssingletest_cache_copyoutput_mismatch7/7(line 670) AssertionError: Lists differ: ['You are a helpful assistant. Help me to [390 chars] is'] != ["You are a helpful assistant. Help me to [385 chars] is']
utilssingletest_dynamic_cache_hardoutput_mismatch7/7(line 522) AssertionError: "Here[57 chars]ave fur, they have four legs, they have a tail[1045 chars]have" != "Here[57 chars]ave four legs, they have a tail, they have a f[1078 chars]They"
video_llavamultitest_small_model_integration_test_llamaoutput_mismatch7/7(line 491) AssertionError: 'USER: \nDescribe the video in details. A[572 chars]ion.' != "USER: \nDescribe the video in details. A[675 chars]ing."
video_llavasingletest_small_model_integration_test_llamaoutput_mismatch7/7(line 491) AssertionError: 'USER: \nDescribe the video in details. A[572 chars]ion.' != "USER: \nDescribe the video in details. A[675 chars]ing."
videomaemultitest_inference_for_pretrainingoutput_mismatch7/7(line 482) AssertionError: The values for attribute 'shape' do not match: torch.Size([]) != torch.Size([1]).
videomaesingletest_inference_for_pretrainingoutput_mismatch7/7(line 482) AssertionError: The values for attribute 'shape' do not match: torch.Size([]) != torch.Size([1]).
viltmultitest_inference_masked_lmoutput_mismatch7/7(line 573) AssertionError: Tensor-likes are not close!
viltsingletest_inference_masked_lmoutput_mismatch7/7(line 573) AssertionError: Tensor-likes are not close!
vitsmultitest_forward_fp16output_mismatch7/7(line 409) AssertionError: Tensor-likes are not close!
vitssingletest_forward_fp16output_mismatch7/7(line 409) AssertionError: Tensor-likes are not close!
zambamultitest_simple_generateoutput_mismatch7/7(line 475) AssertionError: The values for attribute 'dtype' do not match: torch.bfloat16 != torch.float32.

Unpinned failure modes

modecount
output_mismatch232
other64
load_error32
OOM20
cuda_runtime12
import_or_config4

Per-model breakdown (all failures)

modelfailuresgpumode mix
generation18multi/singlecuda_runtime 12 output_mismatch 3 other 3
minimax_m3_vl12multi/singleload_error 12
deepseek_v328multi/singleload_error 8
diffusion_gemma8multi/singleoutput_mismatch 8
grounding_dino8multi/singleoutput_mismatch 8
moshi8multi/singleoutput_mismatch 8
recurrent_gemma8multi/singleoutput_mismatch 6 other 2
peft_integration8multi/singleother 8
cohere2_vision7multi/singleother 4 OOM 2 output_mismatch 1
qwen3_vl_moe7multiother 6 OOM 1
utils7multi/singleoutput_mismatch 4 other 3
bloom6multi/singleoutput_mismatch 6
bridgetower6multi/singleother 6
janus6multi/singleother 6
kosmos26multi/singleoutput_mismatch 6
layoutlmv26multi/singleoutput_mismatch 6
mluke6multi/singleoutput_mismatch 6
mm_grounding_dino6multi/singleoutput_mismatch 6
plbart6multi/singleoutput_mismatch 6
seamless_m4t6multi/singleother 6
starcoder26multi/singleoutput_mismatch 6
blip_25multi/singleoutput_mismatch 5
deepseek_vl_hybrid5multi/singleother 2 OOM 2 output_mismatch 1
bamba4multi/singleoutput_mismatch 4
blt4multi/singleimport_or_config 4
colqwen24multi/singleother 2 output_mismatch 2
deepseek_v44multi/singleload_error 4
exaone44multi/singleoutput_mismatch 3 OOM 1
kimi_k254multi/singleother 2 output_mismatch 2
luke4multi/singleoutput_mismatch 4
lw_detr4multi/singleoutput_mismatch 4
mamba24multi/singleOOM 4
mimi4multi/singleoutput_mismatch 4
minicpm34multi/singleother 4
ministral34multi/singleoutput_mismatch 4
mistral44multi/singleload_error 4
mixtral4multi/singleoutput_mismatch 4
mllama4multi/singleoutput_mismatch 4
musicgen4multi/singleoutput_mismatch 4
oneformer4multi/singleoutput_mismatch 4
pvt4multi/singleoutput_mismatch 4
pvt_v24multi/singleoutput_mismatch 4
qwen2_moe4multi/singleoutput_mismatch 2 other 2
qwen34multi/singleoutput_mismatch 4
qwen3_54multi/singleoutput_mismatch 4
qwen3_moe4multiload_error 4
rag4singleoutput_mismatch 4
smollm34multi/singleoutput_mismatch 4
zamba4multi/singleOOM 3 output_mismatch 1
llama43multi/singleOOM 3
big_bird2multi/singleoutput_mismatch 2
bitnet2multi/singleoutput_mismatch 2
cohere22multi/singleoutput_mismatch 2
convnextv22multi/singleoutput_mismatch 2
cvt2multi/singleoutput_mismatch 2
cwm2multiOOM 2
dab_detr2multi/singleoutput_mismatch 2
deepseek_vl2multi/singleoutput_mismatch 2
dia2multi/singleoutput_mismatch 2
diffllama2multi/singleoutput_mismatch 2
efficientnet2multi/singleoutput_mismatch 2
eomt_dinov32multi/singleoutput_mismatch 2
exaone_moe2multi/singleoutput_mismatch 2
fastspeech2_conformer2multi/singleoutput_mismatch 2
fuyu2multi/singleoutput_mismatch 2
glm4_moe_lite2multi/singleOOM 2
glm_image2multi/singleother 2
granite2multi/singleoutput_mismatch 2
hrm_text2multi/singleoutput_mismatch 2
hy_v32multi/singleoutput_mismatch 2
jais22multi/singleoutput_mismatch 2
jamba2multi/singleoutput_mismatch 2
kosmos2_52multi/singleoutput_mismatch 2
llava_next_video2multi/singleoutput_mismatch 2
m2m_1002multi/singleoutput_mismatch 2
minicpmv4_62multi/singleother 2
ministral2multi/singleoutput_mismatch 2
nemotron2multi/singleoutput_mismatch 2
nllb_moe2multi/singleoutput_mismatch 2
ovis22multi/singleoutput_mismatch 2
pegasus2multi/singleother 2
persimmon2multi/singleoutput_mismatch 2
phi32multi/singleother 2
reformer2multi/singleoutput_mismatch 2
regnet2multi/singleoutput_mismatch 2
seamless_m4t_v22multi/singleoutput_mismatch 2
superpoint2multi/singleoutput_mismatch 2
swinv22multi/singleoutput_mismatch 2
t52multi/singleoutput_mismatch 2
univnet2multi/singleoutput_mismatch 2
video_llava2multi/singleoutput_mismatch 2
videomae2multi/singleoutput_mismatch 2
vilt2multi/singleoutput_mismatch 2
vits2multi/singleoutput_mismatch 2

Pinned clusters (CI bisect)

(none)

Flaky (CI flagged)

(none)

Unpinned — samples per mode

These failures persisted across the window but CI couldn't attribute a bad commit. They likely regressed before the 7-day bisect window. Showing the most-recently-seen samples per failure mode.

output_mismatch 232 unpinned failures — sample of 5

modelgputestdaystrace excerpt
zambamultitest_simple_generate7/7(line 475) AssertionError: The values for attribute 'dtype' do not match: torch.bfloat16 != torch.float32.
vitsmultitest_forward_fp167/7(line 409) AssertionError: Tensor-likes are not close!
vitssingletest_forward_fp167/7(line 409) AssertionError: Tensor-likes are not close!
viltmultitest_inference_masked_lm7/7(line 573) AssertionError: Tensor-likes are not close!
viltsingletest_inference_masked_lm7/7(line 573) AssertionError: Tensor-likes are not close!

OOM 20 unpinned failures — sample of 5

modelgputestdaystrace excerpt
zambamultitest_simple_batched_generate_with_padding7/7(line 1374) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 228.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 12.69 MiB is free. Process 612705 has 22.28 GiB memory in use. Of the allocated mem…
zambasingletest_simple_batched_generate_with_padding7/7(line 1374) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 228.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 12.69 MiB is free. Process 835377 has 22.28 GiB memory in use. Of the allocated mem…
zambasingletest_simple_generate7/7(line 1374) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 228.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 12.69 MiB is free. Process 835377 has 22.28 GiB memory in use. Of the allocated mem…
qwen3_vl_moemultitest_small_model_integration_test_with_video7/7(line 1240) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 768.00 MiB. GPU 1 has a total capacity of 22.30 GiB of which 358.69 MiB is free. Process 127736 has 21.95 GiB memory in use. Of the allocated me…
mamba2singletest_simple_generate7/7(line 1374) torch.OutOfMemoryError: CUDA out of memory. Tried to allocate 64.00 MiB. GPU 0 has a total capacity of 22.30 GiB of which 28.69 MiB is free. Process 618194 has 22.27 GiB memory in use. Of the allocated memo…

load_error 32 unpinned failures — sample of 5

modelgputestdaystrace excerpt
qwen3_moemultitest_model_15b_a2b_generation7/7(line 74) ValueError: Some modules are dispatched on the CPU or the disk. Make sure you have enough GPU RAM to fit the quantized model. If you want to dispatch the model on the CPU or the disk while keeping these modul…
qwen3_moemultitest_model_15b_a2b_logits7/7(line 74) ValueError: Some modules are dispatched on the CPU or the disk. Make sure you have enough GPU RAM to fit the quantized model. If you want to dispatch the model on the CPU or the disk while keeping these modul…
qwen3_moemultitest_model_15b_a2b_long_prompt_sdpa7/7(line 74) ValueError: Some modules are dispatched on the CPU or the disk. Make sure you have enough GPU RAM to fit the quantized model. If you want to dispatch the model on the CPU or the disk while keeping these modul…
qwen3_moemultitest_speculative_generation7/7(line 74) ValueError: Some modules are dispatched on the CPU or the disk. Make sure you have enough GPU RAM to fit the quantized model. If you want to dispatch the model on the CPU or the disk while keeping these modul…
mistral4multitest_mistral_small_4_generation7/7(line 531) ValueError: The current `device_map` had weights offloaded to the disk, which needed to be re-saved. This is either because the weights are not in `safetensors` format, or because the model uses an internal …

cuda_runtime 12 unpinned failures — sample of 5

modelgputestdaystrace excerpt
generationsingletest_validate_assistant7/7(line 1920) torch.AcceleratorError: CUDA error: device-side assert triggered
generationsingletest_matches_immediate_check_at_max_length7/7(line 1374) torch.AcceleratorError: CUDA error: device-side assert triggered
generationsingletest_matches_immediate_check_on_a_batch7/7(line 1374) torch.AcceleratorError: CUDA error: device-side assert triggered
generationsingletest_matches_immediate_check_when_stopping_early7/7(line 1374) torch.AcceleratorError: CUDA error: device-side assert triggered
generationsingletest_matches_immediate_check_with_a_sliding_window7/7(line 1374) torch.AcceleratorError: CUDA error: device-side assert triggered

import_or_config 4 unpinned failures — sample of 4

modelgputestdaystrace excerpt
bltmultitest_model_logits7/7(line 312) AttributeError: 'BltConfig' object has no attribute 'num_hidden_layers'
bltmultitest_model_logits_bf167/7(line 312) AttributeError: 'BltConfig' object has no attribute 'num_hidden_layers'
bltsingletest_model_logits7/7(line 312) AttributeError: 'BltConfig' object has no attribute 'num_hidden_layers'
bltsingletest_model_logits_bf167/7(line 312) AttributeError: 'BltConfig' object has no attribute 'num_hidden_layers'

other 64 unpinned failures — sample of 5

modelgputestdaystrace excerpt
utilssingletest_chunked_prefill_initializes_static_cache_eagerly7/7(line 331) TypeError: Pointer argument must be either uint64 or have data_ptr method
utilsmultitest_chunked_prefill_initializes_static_cache_eagerly7/7(line 331) TypeError: Pointer argument must be either uint64 or have data_ptr method
utilsmultitest_data_parallel_dynamic_cache7/7(line 1834) RuntimeError: a Tensor with 2 elements cannot be converted to Scalar
seamless_m4tmultitest_speech_to_speech_model7/7(line 421) ValueError: Invalid input type. Must be a single audio or a list of audio
seamless_m4tmultitest_speech_to_text_model7/7(line 421) ValueError: Invalid input type. Must be a single audio or a list of audio