Aller directement au contenu
  • Vaultwarden (update)

    Announcements
    5
    3
    0 Votes
    5 Messages
    87 Vues
    fariasF
    Mise à jours en cours : [image: image.jpeg]
  • NodeBB (update)

    Announcements update nodebb
    36
    0 Votes
    36 Messages
    998 Vues
    fariasF
    Mise à jours. # git reset --hard v4.16.0
  • Blocage du jours

    Scan Attack
    71
    0 Votes
    71 Messages
    2k Vues
    fariasF
    # grep "wp-login.php" /var/log/apache2/access.*.log | sed 's/:/ /g' | awk '{print $2}' | sort -n | uniq -c | sort -n | tail -10 1 90.116.27.237 1 92.180.239.91 1 94.31.70.217 2 174.138.77.239 2 185.19.40.76 2 188.130.129.62 2 45.133.112.95 2 91.132.124.66 13 87.232.150.147 489 150.230.106.64
  • Configuration MailCow : PTR Record/HELO Check

    Linux
    12
    3
    0 Votes
    12 Messages
    218 Vues
    fariasF
    Logs : # telnet smtp.arias-frederic.org 25 Trying 80.15.48.50... Connected to smtp.arias-frederic.org. Escape character is '^]'. 220-mail.arias-frederic.org ESMTP Postcow 220 mail.arias-frederic.org ESMTP Postcow HELO testing.com 250 mail.arias-frederic.org EHLO testing.com 250-mail.arias-frederic.org 250-SIZE 104857600 250-ETRN 250-STARTTLS 250-ENHANCEDSTATUSCODES 250 8BITMIME
  • Proxmox Balkany (update)

    Announcements proxmox update
    37
    0 Votes
    37 Messages
    1k Vues
    fariasF
    Update : [image: image.jpeg]
  • LLama-swap : Temps de compression

    Déplacé IA llama-swap
    5
    2
    0 Votes
    5 Messages
    107 Vues
    Tuxedo17T
    Update : [image: llama-swap-compression-hourly.png]
  • VM PixelFed (Update)

    Announcements update pixelfed
    6
    0 Votes
    6 Messages
    177 Vues
    fariasF
    Update de composer : # composer self-update Upgrading to version 2.10.3 (stable channel). Use composer self-update --rollback to return to version 2.9.5 la mise à jours : # cd /www/pixelfed/pixelfed # systemctl stop pixelfed # composer install --no-ansi --no-dev --no-interaction --no-progress --no-scripts --optimize-autoloader Composer plugins have been disabled for safety in this non-interactive session. Set COMPOSER_ALLOW_SUPERUSER=1 if you want to allow plugins to run as root/super user. Do not run Composer as root/super user! See https://getcomposer.org/root for details Installing dependencies from lock file Verifying lock file contents can be installed on current platform. Nothing to install, update or remove Generating optimized autoload files Class App\Rules\WebFinger located in ./app/Rules/Webfinger.php does not comply with psr-4 autoloading standard (rule: App\ => ./app). Skipping. 85 packages you are using are looking for funding. Use the `composer fund` command to find out more! root@insta:/www/pixelfed/pixelfed# composer dump-autoload --optimize Do not run Composer as root/super user! See https://getcomposer.org/root for details Continue as root/super user [yes]? yes Generating optimized autoload files Class App\Rules\WebFinger located in ./app/Rules/Webfinger.php does not comply with psr-4 autoloading standard (rule: App\ => ./app). Skipping. > Illuminate\Foundation\ComposerScripts::postAutoloadDump > @php artisan package:discover --ansi INFO Discovering packages. buzz/laravel-h-captcha ...................................................................................................................... DONE intervention/image-laravel .................................................................................................................. DONE jenssegers/agent ............................................................................................................................ DONE laravel-notification-channels/expo .......................................................................................................... DONE laravel-notification-channels/webpush ....................................................................................................... DONE laravel/horizon ............................................................................................................................. DONE laravel/pulse ............................................................................................................................... DONE laravel/tinker .............................................................................................................................. DONE laravel/ui .................................................................................................................................. DONE livewire/livewire ........................................................................................................................... DONE nesbot/carbon ............................................................................................................................... DONE nunomaduro/termwind ......................................................................................................................... DONE pbmedia/laravel-ffmpeg ...................................................................................................................... DONE pixelfed/laravel-snowflake .................................................................................................................. DONE spatie/laravel-backup ....................................................................................................................... DONE spatie/laravel-image-optimizer .............................................................................................................. DONE spatie/laravel-signal-aware-command ......................................................................................................... DONE stevebauman/purify .......................................................................................................................... DONE Generated optimized autoload files containing 10045 classes # php artisan config:cache INFO Configuration cached successfully. # php artisan route:cache INFO Routes cached successfully. # php artisan migrate --force INFO Nothing to migrate. # chown -R pixelfed:pixelfed /www/pixelfed/pixelfed/ # systemctl start pixelfed
  • Update GO-LANG

    Linux
    1
    0 Votes
    1 Messages
    22 Vues
    Personne n'a répondu
  • Crash LLama.cpp

    IA
    1
    0 Votes
    1 Messages
    21 Vues
    Personne n'a répondu
  • Installation Driver 610.57.04

    IA
    3
    1
    0 Votes
    3 Messages
    49 Vues
    Tuxedo17T
    Pas de crash : # nvidia-smi Tue Sep 1 17:35:06 2026 +-----------------------------------------------------------------------------------------+ | NVIDIA-SMI 610.57.04 KMD Version: 610.57.04 CUDA UMD Version: 13.3 | +-----------------------------------------+------------------------+----------------------+ | GPU Name Persistence-M | Bus-Id Disp.A | Volatile Uncorr. ECC | | Fan Temp Perf Pwr:Usage/Cap | Memory-Usage | GPU-Util Compute M. | | | | MIG M. | |=========================================+========================+======================| | 0 NVIDIA GeForce RTX 3060 ... Off | 00000000:01:00.0 Off | N/A | | N/A 63C P0 36W / 115W | 2096MiB / 6144MiB | 12% Default | | | | N/A | +-----------------------------------------+------------------------+----------------------+ | 1 NVIDIA GeForce RTX 5060 Ti Off | 00000000:05:00.0 Off | N/A | | 30% 51C P1 49W / 180W | 7894MiB / 16311MiB | 32% Default | | | | N/A | +-----------------------------------------+------------------------+----------------------+ +-----------------------------------------------------------------------------------------+ | Processes: | | GPU GI CI PID Type Process name GPU Memory | | ID ID Usage | |=========================================================================================| | 0 N/A N/A 3587 C /usr/local/bin/llama-server 2088MiB | | 1 N/A N/A 3587 C /usr/local/bin/llama-server 7886MiB | +-----------------------------------------------------------------------------------------+
  • Crash Tuxedo :

    IA
    5
    0 Votes
    5 Messages
    71 Vues
    Tuxedo17T
    C’est mieux : [image: 7f991335-22e6-424f-9fe4-4495751d1e66-image.jpeg]
  • garmin-to-fittrackee (install)

    Announcements
    5
    0 Votes
    5 Messages
    56 Vues
    fariasF
    C’est bon cela fonctionne à nouveau … ERROR [sports/get_fittrackee_sport_by_garmin_id] Not Fittrackee match with garmin sport id 240 sports.py:134 ERROR [sports/get_fittrackee_sport_by_garmin_id] Set to cycling road sports.py:137 INFO [main/_fetch_garmin_activity_file] Activity data downloaded to file /tmp/24145714013.zip main.py:241 [08/31/26 13:35:04] INFO [fittrackee/upload_workout] Activity added on Fittrackee with id KdjFEeZJPpk24idGwDRfSS fittrackee.py:288 INFO [main/sync] Fetching activities on Garminfrom Aug 30, 2026to Aug 31, 2026 main.py:163 [08/31/26 13:35:05] INFO [main/_fetch_garmin_activity_file] Activity data downloaded to file /tmp/24181394659.zip main.py:241 [08/31/26 13:35:20] INFO [fittrackee/upload_workout] Activity added on Fittrackee with id G8EdcT5tnX29hNE7ucuwte fittrackee.py:288 Le surf est vue comme du vélo … misère.
  • FitTrackee (update)

    Announcements fittrackee update
    14
    0 Votes
    14 Messages
    386 Vues
    fariasF
    Update done : $ ftcli db upgrade INFO [alembic.runtime.migration] Context impl PostgresqlImpl. INFO [alembic.runtime.migration] Will assume transactional DDL. INFO [alembic.runtime.migration] Running upgrade 3d0c336b7a9e -> 2e3a59ebbc59, update blacklisted_tokens table
  • llama.cpp : Impact de "-fa on" sur le Prompt Speed

    Déplacé IA llama.cpp
    3
    1
    1 Votes
    3 Messages
    97 Vues
    Tuxedo17T
    Je viens de refaire un graphique : [image: llama-swap-prompt-speed-daily.png]
  • PC IA sous Ubuntu

    DSI recherche
    7
    0 Votes
    7 Messages
    145 Vues
    Tuxedo17T
    Voici le panier : https://ldlc.com/s/3PYTTLK
  • llama-swap & demeter-sante.fr : Calcul du coût.

    IA
    1
    2
    1 Votes
    1 Messages
    36 Vues
    Personne n'a répondu
  • Ressource pour Dawarich

    Dawarich dawarich ressource
    1
    7
    0 Votes
    1 Messages
    40 Vues
    Personne n'a répondu
  • Tuxedo 17 : Installation vllm

    Déplacé IA vllm
    4
    0 Votes
    4 Messages
    112 Vues
    Tuxedo17T
    Misère … # /root/vllm-env/bin/vllm serve nvidia/Qwen3.6-35B-A3B-NVFP4 --port 8002 --gpu-memory-utilization 0.95 --max-model-len 180000 --max-num-seqs 1 --kv-cache-dtype fp8 --enable-chunked-prefill --max-num-batched-tokens 4096 --attention-backend flashinfer --reasoning-parser qwen3 WARNING 08-07 13:29:38 [cuda.py:959] Detected different devices in the system: NVIDIA GeForce RTX 3060 Laptop GPU, NVIDIA GeForce RTX 5060 Ti. Please make sure to set `CUDA_DEVICE_ORDER=PCI_BUS_ID` to avoid unexpected behavior. (APIServer pid=1833126) INFO 08-07 13:29:49 [api_utils.py:345] (APIServer pid=1833126) INFO 08-07 13:29:49 [api_utils.py:345] █ █ █▄ ▄█ (APIServer pid=1833126) INFO 08-07 13:29:49 [api_utils.py:345] ▄▄ ▄█ █ █ █ ▀▄▀ █ version 0.26.0 (APIServer pid=1833126) INFO 08-07 13:29:49 [api_utils.py:345] █▄█▀ █ █ █ █ model nvidia/Qwen3.6-35B-A3B-NVFP4 (APIServer pid=1833126) INFO 08-07 13:29:49 [api_utils.py:345] ▀▀ ▀▀▀▀▀ ▀▀▀▀▀ ▀ ▀ (APIServer pid=1833126) INFO 08-07 13:29:49 [api_utils.py:345] (APIServer pid=1833126) INFO 08-07 13:29:49 [api_utils.py:273] non-default args: {'model_tag': 'nvidia/Qwen3.6-35B-A3B-NVFP4', 'port': 8002, 'model': 'nvidia/Qwen3.6-35B-A3B-NVFP4', 'max_model_len': 180000, 'attention_backend': 'flashinfer', 'reasoning_parser': 'qwen3', 'gpu_memory_utilization': 0.95, 'kv_cache_dtype': 'fp8', 'max_num_batched_tokens': 4096, 'max_num_seqs': 1, 'enable_chunked_prefill': True} (APIServer pid=1833126) Warning: You are sending unauthenticated requests to the HF Hub. Please set a HF_TOKEN to enable higher rate limits and faster downloads. config.json: 100%|████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████| 58.1k/58.1k [00:00<00:00, 93.7MB/s] preprocessor_config.json: 100%|███████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████| 390/390 [00:00<00:00, 2.29MB/s] (APIServer pid=1833126) INFO 08-07 13:30:04 [model.py:623] Resolved architecture: Qwen3_5MoeForConditionalGeneration (APIServer pid=1833126) INFO 08-07 13:30:04 [model.py:1788] Using max model len 180000 (APIServer pid=1833126) INFO 08-07 13:30:05 [cache.py:285] Using fp8 data type to store kv cache. It reduces the GPU memory footprint and boosts the performance. Meanwhile, it may cause accuracy drop without a proper scaling factor (APIServer pid=1833126) INFO 08-07 13:30:05 [scheduler.py:252] Chunked prefill is enabled with max_num_batched_tokens=4096. (APIServer pid=1833126) WARNING 08-07 13:30:05 [modelopt.py:385] Detected ModelOpt fp8 checkpoint (quant_algo=FP8). Please note that the format is experimental and could change. (APIServer pid=1833126) WARNING 08-07 13:30:05 [modelopt.py:1034] Detected ModelOpt NVFP4 checkpoint (quant_algo=NVFP4). Please note that the format is experimental and could change in future. (APIServer pid=1833126) WARNING 08-07 13:30:05 [modelopt.py:1034] Detected ModelOpt NVFP4 checkpoint (quant_algo=W4A16_NVFP4). Please note that the format is experimental and could change in future. (APIServer pid=1833126) WARNING 08-07 13:30:05 [modelopt.py:1707] Detected ModelOpt MXFP8 checkpoint. Please note that the format is experimental and could change in future. (APIServer pid=1833126) INFO 08-07 13:30:05 [vllm.py:1109] Asynchronous scheduling is enabled. (APIServer pid=1833126) INFO 08-07 13:30:05 [kernel.py:295] Final IR op priority after setting platform defaults: IrOpPriorityConfig(rms_norm=['native'], fused_add_rms_norm=['native']) tokenizer_config.json: 100%|██████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████| 16.7k/16.7k [00:00<00:00, 26.1MB/s] vocab.json: 100%|█████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████| 6.72M/6.72M [00:00<00:00, 32.2MB/s] tokenizer.json: downloading bytes: ████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████| 5.67MB, 532kB/s tokenizer.json: reconstructing file: 100%|████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████| 12.8MB / 12.8MB, 1.22MB/s chat_template.jinja: 100%|████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████| 7.76k/7.76k [00:00<00:00, 23.5MB/s] generation_config.json: 100%|██████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████| 202/202 [00:00<00:00, 998kB/s] video_preprocessor_config.json: 100%|█████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████████| 385/385 [00:00<00:00, 1.41MB/s] (APIServer pid=1833126) [transformers] The `use_fast` parameter is deprecated and will be removed in a future version. Use `backend="torchvision"` instead of `use_fast=True`, or `backend="pil"` instead of `use_fast=False`. WARNING 08-07 13:30:34 [cuda.py:959] Detected different devices in the system: NVIDIA GeForce RTX 3060 Laptop GPU, NVIDIA GeForce RTX 5060 Ti. Please make sure to set `CUDA_DEVICE_ORDER=PCI_BUS_ID` to avoid unexpected behavior. (EngineCore pid=1833533) INFO 08-07 13:30:41 [core.py:116] Initializing a V1 LLM engine (v0.26.0) with config: model='nvidia/Qwen3.6-35B-A3B-NVFP4', speculative_config=None, tokenizer='nvidia/Qwen3.6-35B-A3B-NVFP4', skip_tokenizer_init=False, tokenizer_mode=auto, revision=None, tokenizer_revision=None, trust_remote_code=False, dtype=torch.bfloat16, max_seq_len=180000, download_dir=None, load_format=auto, tensor_parallel_size=1, pipeline_parallel_size=1, data_parallel_size=1, decode_context_parallel_size=1, dcp_comm_backend=ag_rs, disable_custom_all_reduce=False, quantization=modelopt_mixed, quantization_config=None, enforce_eager=False, enable_return_routed_experts=False, kv_cache_dtype=fp8, device_config=cuda, structured_outputs_config=StructuredOutputsConfig(backend='auto', disable_any_whitespace=False, disable_additional_properties=False, reasoning_parser='qwen3', reasoning_parser_plugin='', enable_in_reasoning=False), observability_config=ObservabilityConfig(show_hidden_metrics_for_version=None, otlp_traces_endpoint=None, collect_detailed_traces=None, kv_cache_metrics=False, kv_cache_metrics_sample=0.01, cudagraph_metrics=False, enable_layerwise_nvtx_tracing=False, enable_mfu_metrics=False, enable_mm_processor_stats=False, enable_logging_iteration_details=False, jit_monitor_mode='warn', jit_monitor_verbose=False), seed=0, served_model_name=nvidia/Qwen3.6-35B-A3B-NVFP4, enable_prefix_caching=False, enable_chunked_prefill=True, pooler_config=None, compilation_config={'mode': <CompilationMode.VLLM_COMPILE: 3>, 'debug_dump_path': None, 'cache_dir': '', 'compile_cache_save_format': 'binary', 'backend': 'inductor', 'custom_ops': ['none'], 'ir_enable_torch_wrap': True, 'splitting_ops': ['vllm::unified_attention_with_output', 'vllm::unified_mla_attention_with_output', 'vllm::mamba_mixer2', 'vllm::mamba_mixer', 'vllm::short_conv', 'vllm::linear_attention', 'vllm::plamo2_mamba_mixer', 'vllm::qwen_gdn_attention_core', 'vllm::gdn_attention_core_xpu', 'vllm::olmo_hybrid_gdn_full_forward', 'vllm::kda_attention', 'vllm::sparse_attn_indexer', 'vllm::rocm_aiter_sparse_attn_indexer', 'vllm::deepseek_v4_attention', 'vllm::hpc_rope_norm_forward', 'vllm::unified_kv_cache_update', 'vllm::unified_mla_kv_cache_update'], 'compile_mm_encoder': False, 'cudagraph_mm_encoder': False, 'encoder_cudagraph_token_budgets': [], 'encoder_cudagraph_max_vision_items_per_batch': 0, 'encoder_cudagraph_max_frames_per_batch': None, 'compile_sizes': [], 'compile_ranges_endpoints': [4096], 'inductor_compile_config': {'enable_auto_functionalized_v2': False, 'size_asserts': False, 'alignment_asserts': False, 'scalar_asserts': False, 'combo_kernels': True, 'benchmark_combo_kernel': True}, 'inductor_passes': {}, 'cudagraph_mode': <CUDAGraphMode.FULL_AND_PIECEWISE: (2, 1)>, 'cudagraph_num_of_warmups': 1, 'cudagraph_capture_sizes': [1, 2], 'cudagraph_copy_inputs': False, 'cudagraph_specialize_lora': True, 'use_inductor_graph_partition': False, 'pass_config': {'fuse_norm_quant': False, 'fuse_act_quant': False, 'fuse_attn_quant': False, 'enable_sp': False, 'fuse_gemm_comms': False, 'fuse_allreduce_rms': False, 'enable_qk_norm_rope_fusion': False, 'fuse_rope_kvcache_cat_mla': False, 'fuse_act_padding': False, 'fuse_qk_norm_rope_kvcache': False}, 'max_cudagraph_capture_size': 2, 'dynamic_shapes_config': {'type': <DynamicShapesType.BACKED: 'backed'>, 'evaluate_guards': False, 'assume_32_bit_indexing': False}, 'local_cache_dir': None, 'fast_moe_cold_start': False, 'static_all_moe_layers': []}, kernel_config=KernelConfig(ir_op_priority=IrOpPriorityConfig(rms_norm=['native'], fused_add_rms_norm=['native']), enable_flashinfer_autotune=True, enable_cutedsl_warmup=True, enable_bf16x3_router_gemm=False, moe_backend='auto', linear_backend='auto') (EngineCore pid=1833533) Warning: You are sending unauthenticated requests to the HF Hub. Please set a HF_TOKEN to enable higher rate limits and faster downloads. (EngineCore pid=1833533) INFO 08-07 13:30:45 [parallel_state.py:1615] world_size=1 rank=0 local_rank=0 distributed_init_method=tcp://10.3.2.87:52967 backend=nccl (EngineCore pid=1833533) INFO 08-07 13:30:45 [parallel_state.py:1946] rank 0 in world size 1 is assigned as DP rank 0, PP rank 0, PCP rank 0, TP rank 0, EP rank 0, EPLB rank N/A (EngineCore pid=1833533) Failed to get device capability: SM 12.x requires CUDA >= 12.9. (EngineCore pid=1833533) Failed to get device capability: SM 12.x requires CUDA >= 12.9. (EngineCore pid=1833533) INFO 08-07 13:30:49 [topk_topp_sampler.py:55] Using FlashInfer for top-p & top-k sampling. (EngineCore pid=1833533) [transformers] The `use_fast` parameter is deprecated and will be removed in a future version. Use `backend="torchvision"` instead of `use_fast=True`, or `backend="pil"` instead of `use_fast=False`. (EngineCore pid=1833533) INFO 08-07 13:31:05 [gpu_model_runner.py:5250] Starting to load model nvidia/Qwen3.6-35B-A3B-NVFP4... (EngineCore pid=1833533) INFO 08-07 13:31:05 [cuda.py:541] Using backend AttentionBackendEnum.FLASH_ATTN for vit attention (EngineCore pid=1833533) INFO 08-07 13:31:05 [mm_encoder_attention.py:373] Using AttentionBackendEnum.FLASH_ATTN for MMEncoderAttention. (EngineCore pid=1833533) INFO 08-07 13:31:05 [__init__.py:635] Selected MarlinFP8ScaledMMLinearKernel for ModelOptFp8LinearMethod (EngineCore pid=1833533) INFO 08-07 13:31:05 [qwen_gdn_linear_attn.py:150] Using Triton/FLA GDN prefill kernel (requested=auto, head_k_dim=128). (EngineCore pid=1833533) INFO 08-07 13:31:05 [nvfp4.py:285] Using 'MARLIN' NvFp4 MoE backend out of potential backends: ['FLASHINFER_TRTLLM', 'FLASHINFER_CUTEDSL', 'FLASHINFER_CUTEDSL_BATCHED', 'FLASHINFER_CUTLASS', 'VLLM_CUTLASS', 'MARLIN', 'HUMMING', 'EMULATION']. (EngineCore pid=1833533) INFO 08-07 13:31:05 [cuda.py:422] Using AttentionBackendEnum.FLASHINFER backend. (EngineCore pid=1833533) ERROR 08-07 13:31:07 [gpu_model_runner.py:5345] Failed to load model - not enough GPU memory. Try lowering --gpu-memory-utilization to free memory for weights, increasing --tensor-parallel-size, or using --quantization. See https://docs.vllm.ai/en/latest/configuration/conserving_memory/ for more tips. (original error: CUDA out of memory. Tried to allocate 256.00 MiB. GPU 0 has a total capacity of 15.52 GiB of which 68.62 MiB is free. Including non-PyTorch memory, this process has 15.44 GiB memory in use. Of the allocated memory 15.15 GiB is allocated by PyTorch, and 78.45 MiB is reserved by PyTorch but unallocated. If reserved but unallocated memory is large try setting PYTORCH_CUDA_ALLOC_CONF=expandable_segments:True to avoid fragmentation. See documentation for Memory Management (https://docs.pytorch.org/docs/stable/notes/cuda.html#optimizing-memory-usage-with-pytorch-cuda-alloc-conf))
  • LLama-swap : test du NVFP4

    Déplacé IA llama-swap
    7
    0 Votes
    7 Messages
    165 Vues
    Tuxedo17T
    NVIDIA : https://huggingface.co/nvidia/Qwen3.6-35B-A3B-NVFP4 Exemple : models: # Premier modèle : Llama 3 8B - name: "meta-llama/Meta-Llama-3-8B-Instruct" command: > vllm serve meta-llama/Meta-Llama-3-8B-Instruct --port 8001 --gpu-memory-utilization 0.85 ready_url: "http://127.0.0" upstream_url: "http://127.0.0"
  • Installation llama.cpp sous Windows 11 avec Ubuntu 22

    Déplacé IA wsl2 llama.cpp
    30
    1 Votes
    30 Messages
    741 Vues
    R
    Je fait un make install avant puis un nouveau test : root@pcremi:/workspace/llama.cpp/build/bin# ./llama-bench -m /models/Qwen3.6-27B-Q4_K_M.gguf ggml_cuda_init: found 1 ROCm devices (Total VRAM: 16304 MiB): Device 0: AMD Radeon RX 9060 XT, gfx1200 (0x1200), VMM: no, Wave Size: 32, VRAM: 16304 MiB | model | size | params | backend | ngl | test | t/s | | ------------------------------ | ---------: | ---------: | ---------- | --: | --------------: | -------------------: | pid:3402 tid:0x757b61d04280 [CreateContext] fail 11 llama-bench: /home/remia/librocdxg/src/wddm/queue.cpp:267: wsl::thunk::ComputeQueue::ComputeQueue(wsl::thunk::WDDMDevice*, void*, uint64_t, std::atomic<long unsigned int>*, std::atomic<long unsigned int>*, volatile int64_t*, uint32_t, uint32_t, bool): Assertion `ret' failed. Aborted (core dumped) Pas mieux.