16396 lines
1.7 MiB
16396 lines
1.7 MiB
==== STARTING EXPERIMENT: qwen3_14b_coding_fft ====
|
||
Log File: saves/qwen3_14b/coding/fft/qwen3_14b_coding_fft_20260430_161911.log
|
||
HF Hub: https://huggingface.co/KKHYA/qwen3-14b-fft-coding
|
||
Timestamp: 2026-04-30 16:19:11
|
||
=====================================
|
||
[INFO|2026-04-30 16:19:20] llamafactory.launcher:144 >> Initializing 8 distributed tasks at: 127.0.0.1:50981
|
||
W0430 16:19:21.559000 83628 site-packages/torch/distributed/run.py:803]
|
||
W0430 16:19:21.559000 83628 site-packages/torch/distributed/run.py:803] *****************************************
|
||
W0430 16:19:21.559000 83628 site-packages/torch/distributed/run.py:803] Setting OMP_NUM_THREADS environment variable for each process to be 1 in default, to avoid your system being overloaded, please further tune the variable for optimal performance in your application as needed.
|
||
W0430 16:19:21.559000 83628 site-packages/torch/distributed/run.py:803] *****************************************
|
||
warmup_ratio is deprecated and will be removed in v5.2. Use `warmup_steps` instead.
|
||
warmup_ratio is deprecated and will be removed in v5.2. Use `warmup_steps` instead.
|
||
warmup_ratio is deprecated and will be removed in v5.2. Use `warmup_steps` instead.
|
||
warmup_ratio is deprecated and will be removed in v5.2. Use `warmup_steps` instead.
|
||
warmup_ratio is deprecated and will be removed in v5.2. Use `warmup_steps` instead.
|
||
warmup_ratio is deprecated and will be removed in v5.2. Use `warmup_steps` instead.
|
||
warmup_ratio is deprecated and will be removed in v5.2. Use `warmup_steps` instead.
|
||
warmup_ratio is deprecated and will be removed in v5.2. Use `warmup_steps` instead.
|
||
[W430 16:19:34.093488596 ProcessGroupNCCL.cpp:924] Warning: TORCH_NCCL_AVOID_RECORD_STREAMS is the default now, this environment variable is thus deprecated. (function operator())
|
||
[W430 16:19:34.154182617 ProcessGroupNCCL.cpp:924] Warning: TORCH_NCCL_AVOID_RECORD_STREAMS is the default now, this environment variable is thus deprecated. (function operator())
|
||
[W430 16:19:34.168539397 ProcessGroupNCCL.cpp:924] Warning: TORCH_NCCL_AVOID_RECORD_STREAMS is the default now, this environment variable is thus deprecated. (function operator())
|
||
[W430 16:19:34.177012669 ProcessGroupNCCL.cpp:924] Warning: TORCH_NCCL_AVOID_RECORD_STREAMS is the default now, this environment variable is thus deprecated. (function operator())
|
||
[W430 16:19:34.263604490 ProcessGroupNCCL.cpp:924] Warning: TORCH_NCCL_AVOID_RECORD_STREAMS is the default now, this environment variable is thus deprecated. (function operator())
|
||
[W430 16:19:34.285989691 ProcessGroupNCCL.cpp:924] Warning: TORCH_NCCL_AVOID_RECORD_STREAMS is the default now, this environment variable is thus deprecated. (function operator())
|
||
[W430 16:19:34.319602131 ProcessGroupNCCL.cpp:924] Warning: TORCH_NCCL_AVOID_RECORD_STREAMS is the default now, this environment variable is thus deprecated. (function operator())
|
||
[W430 16:19:34.343293697 ProcessGroupNCCL.cpp:924] Warning: TORCH_NCCL_AVOID_RECORD_STREAMS is the default now, this environment variable is thus deprecated. (function operator())
|
||
ywang29-p4d-debug-2-worker-0:83696:83696 [0] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83696:83696 [0] NCCL INFO Bootstrap: Using eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83696:83696 [0] NCCL INFO cudaDriverVersion 13000
|
||
ywang29-p4d-debug-2-worker-0:83696:83696 [0] NCCL INFO NCCL version 2.27.7+cuda13.0
|
||
ywang29-p4d-debug-2-worker-0:83703:83703 [7] NCCL INFO cudaDriverVersion 13000
|
||
ywang29-p4d-debug-2-worker-0:83703:83703 [7] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83703:83703 [7] NCCL INFO Bootstrap: Using eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83703:83703 [7] NCCL INFO NCCL version 2.27.7+cuda13.0
|
||
ywang29-p4d-debug-2-worker-0:83703:83703 [7] NCCL INFO Comm config Blocking set to 1
|
||
ywang29-p4d-debug-2-worker-0:83702:83702 [6] NCCL INFO cudaDriverVersion 13000
|
||
ywang29-p4d-debug-2-worker-0:83702:83702 [6] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83702:83702 [6] NCCL INFO Bootstrap: Using eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83702:83702 [6] NCCL INFO NCCL version 2.27.7+cuda13.0
|
||
ywang29-p4d-debug-2-worker-0:83698:83698 [2] NCCL INFO cudaDriverVersion 13000
|
||
ywang29-p4d-debug-2-worker-0:83698:83698 [2] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83698:83698 [2] NCCL INFO Bootstrap: Using eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83698:83698 [2] NCCL INFO NCCL version 2.27.7+cuda13.0
|
||
ywang29-p4d-debug-2-worker-0:83702:83702 [6] NCCL INFO Comm config Blocking set to 1
|
||
ywang29-p4d-debug-2-worker-0:83698:83698 [2] NCCL INFO Comm config Blocking set to 1
|
||
ywang29-p4d-debug-2-worker-0:83697:83697 [1] NCCL INFO cudaDriverVersion 13000
|
||
ywang29-p4d-debug-2-worker-0:83697:83697 [1] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83697:83697 [1] NCCL INFO Bootstrap: Using eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83697:83697 [1] NCCL INFO NCCL version 2.27.7+cuda13.0
|
||
ywang29-p4d-debug-2-worker-0:83697:83697 [1] NCCL INFO Comm config Blocking set to 1
|
||
ywang29-p4d-debug-2-worker-0:83701:83701 [5] NCCL INFO cudaDriverVersion 13000
|
||
ywang29-p4d-debug-2-worker-0:83701:83701 [5] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83701:83701 [5] NCCL INFO Bootstrap: Using eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83701:83701 [5] NCCL INFO NCCL version 2.27.7+cuda13.0
|
||
ywang29-p4d-debug-2-worker-0:83700:83700 [4] NCCL INFO cudaDriverVersion 13000
|
||
ywang29-p4d-debug-2-worker-0:83699:83699 [3] NCCL INFO cudaDriverVersion 13000
|
||
ywang29-p4d-debug-2-worker-0:83701:83701 [5] NCCL INFO Comm config Blocking set to 1
|
||
ywang29-p4d-debug-2-worker-0:83700:83700 [4] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83699:83699 [3] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83700:83700 [4] NCCL INFO Bootstrap: Using eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83700:83700 [4] NCCL INFO NCCL version 2.27.7+cuda13.0
|
||
ywang29-p4d-debug-2-worker-0:83699:83699 [3] NCCL INFO Bootstrap: Using eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83699:83699 [3] NCCL INFO NCCL version 2.27.7+cuda13.0
|
||
ywang29-p4d-debug-2-worker-0:83699:83699 [3] NCCL INFO Comm config Blocking set to 1
|
||
ywang29-p4d-debug-2-worker-0:83700:83700 [4] NCCL INFO Comm config Blocking set to 1
|
||
ywang29-p4d-debug-2-worker-0:83696:83696 [0] NCCL INFO Comm config Blocking set to 1
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/Plugin: Plugin name set by env to libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/Plugin: Loaded net plugin Libfabric (v10)
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v10 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v9 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v8 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v7 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v6 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO Successfully loaded external plugin libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/OFI Initializing aws-ofi-nccl 1.17.1
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/OFI Using Libfabric version 2.3
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/OFI Using CUDA driver version 13000 with runtime 13000
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/OFI Configuring AWS-specific options
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/OFI Setting provider_filter to efa
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/OFI Running on p4de.24xlarge platform, topology file /opt/amazon/ofi-nccl/share/aws-ofi-nccl/xml/p4de-24xl-topo.xml
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/OFI Internode latency set at 75.0 us
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/OFI Using transport protocol SENDRECV (platform set)
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83703:83977 [7] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):208 NCCL WARN NET/OFI Failed to initialize sendrecv protocol
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83703:83977 [7] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):353 NCCL WARN NET/OFI aws-ofi-nccl initialization failed
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83703:83977 [7] ncclResult_t nccl_net_ofi_init_v6(ncclDebugLogger_t):162 NCCL WARN NET/OFI Initializing plugin failed
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/IB : No device found.
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/IB : Using [RO]; OOB eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/Socket : Using [0]eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO Initialized NET plugin Socket
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO Using network Socket
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO ncclCommInitRankConfig comm 0x55f5a30b2870 rank 7 nranks 8 cudaDev 7 nvmlDev 7 busId a01d0 commId 0xc553500499762690 - Init START
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/Plugin: Plugin name set by env to libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/Plugin: Loaded net plugin Libfabric (v10)
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v10 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v9 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v8 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v7 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v6 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO Successfully loaded external plugin libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/OFI Initializing aws-ofi-nccl 1.17.1
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/OFI Using Libfabric version 2.3
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/OFI Using CUDA driver version 13000 with runtime 13000
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/OFI Configuring AWS-specific options
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/OFI Setting provider_filter to efa
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/OFI Running on p4de.24xlarge platform, topology file /opt/amazon/ofi-nccl/share/aws-ofi-nccl/xml/p4de-24xl-topo.xml
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/OFI Internode latency set at 75.0 us
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/OFI Using transport protocol SENDRECV (platform set)
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83702:83978 [6] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):208 NCCL WARN NET/OFI Failed to initialize sendrecv protocol
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83702:83978 [6] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):353 NCCL WARN NET/OFI aws-ofi-nccl initialization failed
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83702:83978 [6] ncclResult_t nccl_net_ofi_init_v6(ncclDebugLogger_t):162 NCCL WARN NET/OFI Initializing plugin failed
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/IB : No device found.
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/IB : Using [RO]; OOB eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/Socket : Using [0]eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO Initialized NET plugin Socket
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO Using network Socket
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/Plugin: Plugin name set by env to libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/Plugin: Loaded net plugin Libfabric (v10)
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v10 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v9 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v8 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v7 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v6 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO Successfully loaded external plugin libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/OFI Initializing aws-ofi-nccl 1.17.1
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/OFI Using Libfabric version 2.3
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/OFI Using CUDA driver version 13000 with runtime 13000
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/OFI Configuring AWS-specific options
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/OFI Setting provider_filter to efa
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/OFI Running on p4de.24xlarge platform, topology file /opt/amazon/ofi-nccl/share/aws-ofi-nccl/xml/p4de-24xl-topo.xml
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/OFI Internode latency set at 75.0 us
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/OFI Using transport protocol SENDRECV (platform set)
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83698:83979 [2] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):208 NCCL WARN NET/OFI Failed to initialize sendrecv protocol
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83698:83979 [2] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):353 NCCL WARN NET/OFI aws-ofi-nccl initialization failed
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83698:83979 [2] ncclResult_t nccl_net_ofi_init_v6(ncclDebugLogger_t):162 NCCL WARN NET/OFI Initializing plugin failed
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/Plugin: Plugin name set by env to libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/Plugin: Loaded net plugin Libfabric (v10)
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v10 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v9 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v8 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v7 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v6 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO Successfully loaded external plugin libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/OFI Initializing aws-ofi-nccl 1.17.1
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/OFI Using Libfabric version 2.3
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/IB : No device found.
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/IB : Using [RO]; OOB eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/Socket : Using [0]eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO Initialized NET plugin Socket
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO Using network Socket
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/OFI Using CUDA driver version 13000 with runtime 13000
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/OFI Configuring AWS-specific options
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/OFI Setting provider_filter to efa
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/OFI Running on p4de.24xlarge platform, topology file /opt/amazon/ofi-nccl/share/aws-ofi-nccl/xml/p4de-24xl-topo.xml
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/OFI Internode latency set at 75.0 us
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/OFI Using transport protocol SENDRECV (platform set)
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83697:83980 [1] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):208 NCCL WARN NET/OFI Failed to initialize sendrecv protocol
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83697:83980 [1] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):353 NCCL WARN NET/OFI aws-ofi-nccl initialization failed
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83697:83980 [1] ncclResult_t nccl_net_ofi_init_v6(ncclDebugLogger_t):162 NCCL WARN NET/OFI Initializing plugin failed
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/IB : No device found.
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/IB : Using [RO]; OOB eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/Socket : Using [0]eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO Initialized NET plugin Socket
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO Using network Socket
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/Plugin: Plugin name set by env to libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/Plugin: Loaded net plugin Libfabric (v10)
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v10 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v9 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v8 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v7 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v6 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO Successfully loaded external plugin libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/OFI Initializing aws-ofi-nccl 1.17.1
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/OFI Using Libfabric version 2.3
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/OFI Using CUDA driver version 13000 with runtime 13000
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/OFI Configuring AWS-specific options
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/OFI Setting provider_filter to efa
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/OFI Running on p4de.24xlarge platform, topology file /opt/amazon/ofi-nccl/share/aws-ofi-nccl/xml/p4de-24xl-topo.xml
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/OFI Internode latency set at 75.0 us
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/OFI Using transport protocol SENDRECV (platform set)
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83700:83983 [4] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):208 NCCL WARN NET/OFI Failed to initialize sendrecv protocol
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83700:83983 [4] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):353 NCCL WARN NET/OFI aws-ofi-nccl initialization failed
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83700:83983 [4] ncclResult_t nccl_net_ofi_init_v6(ncclDebugLogger_t):162 NCCL WARN NET/OFI Initializing plugin failed
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/IB : No device found.
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/IB : Using [RO]; OOB eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/Socket : Using [0]eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO Initialized NET plugin Socket
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO Using network Socket
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/Plugin: Plugin name set by env to libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/Plugin: Loaded net plugin Libfabric (v10)
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v10 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v9 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v8 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v7 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v6 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Successfully loaded external plugin libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/OFI Initializing aws-ofi-nccl 1.17.1
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/OFI Using Libfabric version 2.3
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/OFI Using CUDA driver version 13000 with runtime 13000
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/OFI Configuring AWS-specific options
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/OFI Setting provider_filter to efa
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/OFI Running on p4de.24xlarge platform, topology file /opt/amazon/ofi-nccl/share/aws-ofi-nccl/xml/p4de-24xl-topo.xml
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/OFI Internode latency set at 75.0 us
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/OFI Using transport protocol SENDRECV (platform set)
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/Plugin: Plugin name set by env to libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/Plugin: Loaded net plugin Libfabric (v10)
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v10 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v9 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v8 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v7 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v6 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO Successfully loaded external plugin libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/OFI Initializing aws-ofi-nccl 1.17.1
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/OFI Using Libfabric version 2.3
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/OFI Using CUDA driver version 13000 with runtime 13000
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/OFI Configuring AWS-specific options
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/OFI Setting provider_filter to efa
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/OFI Running on p4de.24xlarge platform, topology file /opt/amazon/ofi-nccl/share/aws-ofi-nccl/xml/p4de-24xl-topo.xml
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/OFI Internode latency set at 75.0 us
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/OFI Using transport protocol SENDRECV (platform set)
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83696:83984 [0] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):208 NCCL WARN NET/OFI Failed to initialize sendrecv protocol
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83696:83984 [0] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):353 NCCL WARN NET/OFI aws-ofi-nccl initialization failed
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83696:83984 [0] ncclResult_t nccl_net_ofi_init_v6(ncclDebugLogger_t):162 NCCL WARN NET/OFI Initializing plugin failed
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/Plugin: Plugin name set by env to libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/Plugin: Loaded net plugin Libfabric (v10)
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v10 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v9 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v8 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v7 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/Plugin: Failed to find ncclCollNetPlugin_v6 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO Successfully loaded external plugin libnccl-net.so
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/OFI Initializing aws-ofi-nccl 1.17.1
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/OFI Using Libfabric version 2.3
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/IB : No device found.
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/IB : Using [RO]; OOB eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/Socket : Using [0]eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Initialized NET plugin Socket
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Using network Socket
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/OFI Using CUDA driver version 13000 with runtime 13000
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/OFI Configuring AWS-specific options
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/OFI Setting provider_filter to efa
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/OFI Running on p4de.24xlarge platform, topology file /opt/amazon/ofi-nccl/share/aws-ofi-nccl/xml/p4de-24xl-topo.xml
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/OFI Internode latency set at 75.0 us
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/OFI Using transport protocol SENDRECV (platform set)
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO ncclCommInitRankConfig comm 0x55a9d0483a80 rank 6 nranks 8 cudaDev 6 nvmlDev 6 busId a01c0 commId 0xc553500499762690 - Init START
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83701:83981 [5] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):208 NCCL WARN NET/OFI Failed to initialize sendrecv protocol
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83701:83981 [5] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):353 NCCL WARN NET/OFI aws-ofi-nccl initialization failed
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83701:83981 [5] ncclResult_t nccl_net_ofi_init_v6(ncclDebugLogger_t):162 NCCL WARN NET/OFI Initializing plugin failed
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/IB : No device found.
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/IB : Using [RO]; OOB eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/Socket : Using [0]eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO Initialized NET plugin Socket
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO Using network Socket
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83699:83982 [3] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):208 NCCL WARN NET/OFI Failed to initialize sendrecv protocol
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83699:83982 [3] int nccl_net_ofi_create_plugin(nccl_net_ofi_plugin_t**):353 NCCL WARN NET/OFI aws-ofi-nccl initialization failed
|
||
|
||
[2026-04-30 16:19:35] ywang29-p4d-debug-2-worker-0:83699:83982 [3] ncclResult_t nccl_net_ofi_init_v6(ncclDebugLogger_t):162 NCCL WARN NET/OFI Initializing plugin failed
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/IB : No device found.
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/IB : Using [RO]; OOB eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NCCL_SOCKET_IFNAME set by environment to eth
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/Socket : Using [0]eth0:10.200.143.174<0>
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO Initialized NET plugin Socket
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO Using network Socket
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO ncclCommInitRankConfig comm 0x55e5169daea0 rank 2 nranks 8 cudaDev 2 nvmlDev 2 busId 201c0 commId 0xc553500499762690 - Init START
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO ncclCommInitRankConfig comm 0x55ede9ca7de0 rank 1 nranks 8 cudaDev 1 nvmlDev 1 busId 101d0 commId 0xc553500499762690 - Init START
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO ncclCommInitRankConfig comm 0x562c99db2980 rank 4 nranks 8 cudaDev 4 nvmlDev 4 busId 901c0 commId 0xc553500499762690 - Init START
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO ncclCommInitRankConfig comm 0x55ad880448e0 rank 0 nranks 8 cudaDev 0 nvmlDev 0 busId 101c0 commId 0xc553500499762690 - Init START
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO RAS client listening socket at ::1<28028>
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO RAS client listening socket at ::1<28028>
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO RAS client listening socket at ::1<28028>
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO ncclCommInitRankConfig comm 0x55dcff54b080 rank 5 nranks 8 cudaDev 5 nvmlDev 5 busId 901d0 commId 0xc553500499762690 - Init START
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO ncclCommInitRankConfig comm 0x5608cb6aab30 rank 3 nranks 8 cudaDev 3 nvmlDev 3 busId 201d0 commId 0xc553500499762690 - Init START
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO RAS client listening socket at ::1<28028>
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO RAS client listening socket at ::1<28028>
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO RAS client listening socket at ::1<28028>
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO RAS client listening socket at ::1<28028>
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO RAS client listening socket at ::1<28028>
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO Bootstrap timings total 0.046884 (create 0.000040, send 0.000072, recv 0.000157, ring 0.002269, delay 0.000001)
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Bootstrap timings total 0.003056 (create 0.000030, send 0.000076, recv 0.000181, ring 0.002306, delay 0.000001)
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO Bootstrap timings total 0.263699 (create 0.000048, send 0.000165, recv 0.260698, ring 0.002269, delay 0.000001)
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO Bootstrap timings total 0.097396 (create 0.000043, send 0.000087, recv 0.000174, ring 0.000655, delay 0.000001)
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO Bootstrap timings total 0.001407 (create 0.000041, send 0.000066, recv 0.000261, ring 0.000689, delay 0.000001)
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO Bootstrap timings total 0.008058 (create 0.000037, send 0.000070, recv 0.006824, ring 0.000173, delay 0.000001)
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO Bootstrap timings total 0.000836 (create 0.000041, send 0.000104, recv 0.000117, ring 0.000130, delay 0.000001)
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO Bootstrap timings total 0.057812 (create 0.000039, send 0.000078, recv 0.057085, ring 0.000125, delay 0.000001)
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO Setting affinity for GPU 1 to 0-23,48-71
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NVLS multicast support is not available on dev 1 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Setting affinity for GPU 0 to 0-23,48-71
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NVLS multicast support is not available on dev 0 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO Setting affinity for GPU 7 to 24-47,72-95
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NVLS multicast support is not available on dev 7 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO Setting affinity for GPU 6 to 24-47,72-95
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NVLS multicast support is not available on dev 6 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO Setting affinity for GPU 5 to 24-47,72-95
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NVLS multicast support is not available on dev 5 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO Setting affinity for GPU 4 to 24-47,72-95
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NVLS multicast support is not available on dev 4 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO Setting affinity for GPU 2 to 0-23,48-71
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NVLS multicast support is not available on dev 2 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO Setting affinity for GPU 3 to 0-23,48-71
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NVLS multicast support is not available on dev 3 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO comm 0x55e5169daea0 rank 2 nRanks 8 nNodes 1 localRanks 8 localRank 2 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO comm 0x55ede9ca7de0 rank 1 nRanks 8 nNodes 1 localRanks 8 localRank 1 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO comm 0x55ad880448e0 rank 0 nRanks 8 nNodes 1 localRanks 8 localRank 0 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO comm 0x55f5a30b2870 rank 7 nRanks 8 nNodes 1 localRanks 8 localRank 7 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO comm 0x55a9d0483a80 rank 6 nRanks 8 nNodes 1 localRanks 8 localRank 6 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO comm 0x55dcff54b080 rank 5 nRanks 8 nNodes 1 localRanks 8 localRank 5 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO comm 0x5608cb6aab30 rank 3 nRanks 8 nNodes 1 localRanks 8 localRank 3 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 00/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO comm 0x562c99db2980 rank 4 nRanks 8 nNodes 1 localRanks 8 localRank 4 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 01/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 02/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 03/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 04/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 05/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO Trees [0] 2/-1/-1->1->0 [1] 2/-1/-1->1->0 [2] 2/-1/-1->1->0 [3] 2/-1/-1->1->0 [4] 2/-1/-1->1->0 [5] 2/-1/-1->1->0 [6] 2/-1/-1->1->0 [7] 2/-1/-1->1->0 [8] 2/-1/-1->1->0 [9] 2/-1/-1->1->0 [10] 2/-1/-1->1->0 [11] 2/-1/-1->1->0 [12] 2/-1/-1->1->0 [13] 2/-1/-1->1->0 [14] 2/-1/-1->1->0 [15] 2/-1/-1->1->0 [16] 2/-1/-1->1->0 [17] 2/-1/-1->1->0 [18] 2/-1/-1->1->0 [19] 2/-1/-1->1->0 [20] 2/-1/-1->1->0 [21] 2/-1/-1->1->0 [22] 2/-1/-1->1->0 [23] 2/-1/-1->1->0
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 06/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 07/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 08/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO Trees [0] -1/-1/-1->7->6 [1] -1/-1/-1->7->6 [2] -1/-1/-1->7->6 [3] -1/-1/-1->7->6 [4] -1/-1/-1->7->6 [5] -1/-1/-1->7->6 [6] -1/-1/-1->7->6 [7] -1/-1/-1->7->6 [8] -1/-1/-1->7->6 [9] -1/-1/-1->7->6 [10] -1/-1/-1->7->6 [11] -1/-1/-1->7->6 [12] -1/-1/-1->7->6 [13] -1/-1/-1->7->6 [14] -1/-1/-1->7->6 [15] -1/-1/-1->7->6 [16] -1/-1/-1->7->6 [17] -1/-1/-1->7->6 [18] -1/-1/-1->7->6 [19] -1/-1/-1->7->6 [20] -1/-1/-1->7->6 [21] -1/-1/-1->7->6 [22] -1/-1/-1->7->6 [23] -1/-1/-1->7->6
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 09/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 10/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 11/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO Trees [0] 7/-1/-1->6->5 [1] 7/-1/-1->6->5 [2] 7/-1/-1->6->5 [3] 7/-1/-1->6->5 [4] 7/-1/-1->6->5 [5] 7/-1/-1->6->5 [6] 7/-1/-1->6->5 [7] 7/-1/-1->6->5 [8] 7/-1/-1->6->5 [9] 7/-1/-1->6->5 [10] 7/-1/-1->6->5 [11] 7/-1/-1->6->5 [12] 7/-1/-1->6->5 [13] 7/-1/-1->6->5 [14] 7/-1/-1->6->5 [15] 7/-1/-1->6->5 [16] 7/-1/-1->6->5 [17] 7/-1/-1->6->5 [18] 7/-1/-1->6->5 [19] 7/-1/-1->6->5 [20] 7/-1/-1->6->5 [21] 7/-1/-1->6->5 [22] 7/-1/-1->6->5 [23] 7/-1/-1->6->5
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO Trees [0] 6/-1/-1->5->4 [1] 6/-1/-1->5->4 [2] 6/-1/-1->5->4 [3] 6/-1/-1->5->4 [4] 6/-1/-1->5->4 [5] 6/-1/-1->5->4 [6] 6/-1/-1->5->4 [7] 6/-1/-1->5->4 [8] 6/-1/-1->5->4 [9] 6/-1/-1->5->4 [10] 6/-1/-1->5->4 [11] 6/-1/-1->5->4 [12] 6/-1/-1->5->4 [13] 6/-1/-1->5->4 [14] 6/-1/-1->5->4 [15] 6/-1/-1->5->4 [16] 6/-1/-1->5->4 [17] 6/-1/-1->5->4 [18] 6/-1/-1->5->4 [19] 6/-1/-1->5->4 [20] 6/-1/-1->5->4 [21] 6/-1/-1->5->4 [22] 6/-1/-1->5->4 [23] 6/-1/-1->5->4
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 12/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO Trees [0] 3/-1/-1->2->1 [1] 3/-1/-1->2->1 [2] 3/-1/-1->2->1 [3] 3/-1/-1->2->1 [4] 3/-1/-1->2->1 [5] 3/-1/-1->2->1 [6] 3/-1/-1->2->1 [7] 3/-1/-1->2->1 [8] 3/-1/-1->2->1 [9] 3/-1/-1->2->1 [10] 3/-1/-1->2->1 [11] 3/-1/-1->2->1 [12] 3/-1/-1->2->1 [13] 3/-1/-1->2->1 [14] 3/-1/-1->2->1 [15] 3/-1/-1->2->1 [16] 3/-1/-1->2->1 [17] 3/-1/-1->2->1 [18] 3/-1/-1->2->1 [19] 3/-1/-1->2->1 [20] 3/-1/-1->2->1 [21] 3/-1/-1->2->1 [22] 3/-1/-1->2->1 [23] 3/-1/-1->2->1
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 13/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 14/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 15/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 16/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 17/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO Trees [0] 5/-1/-1->4->3 [1] 5/-1/-1->4->3 [2] 5/-1/-1->4->3 [3] 5/-1/-1->4->3 [4] 5/-1/-1->4->3 [5] 5/-1/-1->4->3 [6] 5/-1/-1->4->3 [7] 5/-1/-1->4->3 [8] 5/-1/-1->4->3 [9] 5/-1/-1->4->3 [10] 5/-1/-1->4->3 [11] 5/-1/-1->4->3 [12] 5/-1/-1->4->3 [13] 5/-1/-1->4->3 [14] 5/-1/-1->4->3 [15] 5/-1/-1->4->3 [16] 5/-1/-1->4->3 [17] 5/-1/-1->4->3 [18] 5/-1/-1->4->3 [19] 5/-1/-1->4->3 [20] 5/-1/-1->4->3 [21] 5/-1/-1->4->3 [22] 5/-1/-1->4->3 [23] 5/-1/-1->4->3
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 18/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 19/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 20/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO Trees [0] 4/-1/-1->3->2 [1] 4/-1/-1->3->2 [2] 4/-1/-1->3->2 [3] 4/-1/-1->3->2 [4] 4/-1/-1->3->2 [5] 4/-1/-1->3->2 [6] 4/-1/-1->3->2 [7] 4/-1/-1->3->2 [8] 4/-1/-1->3->2 [9] 4/-1/-1->3->2 [10] 4/-1/-1->3->2 [11] 4/-1/-1->3->2 [12] 4/-1/-1->3->2 [13] 4/-1/-1->3->2 [14] 4/-1/-1->3->2 [15] 4/-1/-1->3->2 [16] 4/-1/-1->3->2 [17] 4/-1/-1->3->2 [18] 4/-1/-1->3->2 [19] 4/-1/-1->3->2 [20] 4/-1/-1->3->2 [21] 4/-1/-1->3->2 [22] 4/-1/-1->3->2 [23] 4/-1/-1->3->2
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 21/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 22/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Channel 23/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Trees [0] 1/-1/-1->0->-1 [1] 1/-1/-1->0->-1 [2] 1/-1/-1->0->-1 [3] 1/-1/-1->0->-1 [4] 1/-1/-1->0->-1 [5] 1/-1/-1->0->-1 [6] 1/-1/-1->0->-1 [7] 1/-1/-1->0->-1 [8] 1/-1/-1->0->-1 [9] 1/-1/-1->0->-1 [10] 1/-1/-1->0->-1 [11] 1/-1/-1->0->-1 [12] 1/-1/-1->0->-1 [13] 1/-1/-1->0->-1 [14] 1/-1/-1->0->-1 [15] 1/-1/-1->0->-1 [16] 1/-1/-1->0->-1 [17] 1/-1/-1->0->-1 [18] 1/-1/-1->0->-1 [19] 1/-1/-1->0->-1 [20] 1/-1/-1->0->-1 [21] 1/-1/-1->0->-1 [22] 1/-1/-1->0->-1 [23] 1/-1/-1->0->-1
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO PROFILER/Plugin: Could not find: libnccl-profiler.so.
|
||
ywang29-p4d-debug-2-worker-0:83702:83994 [6] NCCL INFO [Proxy Service UDS] Device 6 CPU core 83
|
||
ywang29-p4d-debug-2-worker-0:83702:83993 [6] NCCL INFO [Proxy Service] Device 6 CPU core 34
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO PROFILER/Plugin: Could not find: libnccl-profiler.so.
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO PROFILER/Plugin: Could not find: libnccl-profiler.so.
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Check P2P Type isAllDirectP2p 1 directMode 0
|
||
ywang29-p4d-debug-2-worker-0:83696:83995 [0] NCCL INFO [Proxy Service] Device 0 CPU core 7
|
||
ywang29-p4d-debug-2-worker-0:83696:83997 [0] NCCL INFO [Proxy Service UDS] Device 0 CPU core 64
|
||
ywang29-p4d-debug-2-worker-0:83699:83998 [3] NCCL INFO [Proxy Service UDS] Device 3 CPU core 62
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO PROFILER/Plugin: Could not find: libnccl-profiler.so.
|
||
ywang29-p4d-debug-2-worker-0:83697:83999 [1] NCCL INFO [Proxy Service] Device 1 CPU core 1
|
||
ywang29-p4d-debug-2-worker-0:83697:84000 [1] NCCL INFO [Proxy Service UDS] Device 1 CPU core 50
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO PROFILER/Plugin: Could not find: libnccl-profiler.so.
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO PROFILER/Plugin: Could not find: libnccl-profiler.so.
|
||
ywang29-p4d-debug-2-worker-0:83700:84002 [4] NCCL INFO [Proxy Service UDS] Device 4 CPU core 95
|
||
ywang29-p4d-debug-2-worker-0:83700:84001 [4] NCCL INFO [Proxy Service] Device 4 CPU core 90
|
||
ywang29-p4d-debug-2-worker-0:83701:84003 [5] NCCL INFO [Proxy Service] Device 5 CPU core 28
|
||
ywang29-p4d-debug-2-worker-0:83701:84004 [5] NCCL INFO [Proxy Service UDS] Device 5 CPU core 91
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO PROFILER/Plugin: Could not find: libnccl-profiler.so.
|
||
ywang29-p4d-debug-2-worker-0:83703:84005 [7] NCCL INFO [Proxy Service] Device 7 CPU core 36
|
||
ywang29-p4d-debug-2-worker-0:83703:84006 [7] NCCL INFO [Proxy Service UDS] Device 7 CPU core 73
|
||
ywang29-p4d-debug-2-worker-0:83699:83996 [3] NCCL INFO [Proxy Service] Device 3 CPU core 63
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO PROFILER/Plugin: Could not find: libnccl-profiler.so.
|
||
ywang29-p4d-debug-2-worker-0:83698:84007 [2] NCCL INFO [Proxy Service] Device 2 CPU core 3
|
||
ywang29-p4d-debug-2-worker-0:83698:84008 [2] NCCL INFO [Proxy Service UDS] Device 2 CPU core 52
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO CC Off, workFifoBytes 1048576
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO TUNER/Plugin: Could not find: libnccl-tuner.so. Using internal tuner plugin.
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO TUNER/Plugin: Failed to find ncclTunerPlugin_v4 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO TUNER/Plugin: Using tuner plugin nccl_ofi_tuner
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO TUNER/Plugin: Could not find: libnccl-tuner.so. Using internal tuner plugin.
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO ncclCommInitRankConfig comm 0x55a9d0483a80 rank 6 nranks 8 cudaDev 6 nvmlDev 6 busId a01c0 commId 0xc553500499762690 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83702:83978 [6] NCCL INFO Init timings - ncclCommInitRankConfig: rank 6 nranks 8 total 0.47 (kernels 0.18, alloc 0.05, bootstrap 0.10, allgathers 0.01, topo 0.05, graphs 0.00, connections 0.06, rest 0.02)
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO TUNER/Plugin: Failed to find ncclTunerPlugin_v4 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO TUNER/Plugin: Using tuner plugin nccl_ofi_tuner
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO ncclCommInitRankConfig comm 0x562c99db2980 rank 4 nranks 8 cudaDev 4 nvmlDev 4 busId 901c0 commId 0xc553500499762690 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83700:83983 [4] NCCL INFO Init timings - ncclCommInitRankConfig: rank 4 nranks 8 total 0.44 (kernels 0.19, alloc 0.10, bootstrap 0.01, allgathers 0.00, topo 0.05, graphs 0.00, connections 0.06, rest 0.02)
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO TUNER/Plugin: Could not find: libnccl-tuner.so. Using internal tuner plugin.
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO TUNER/Plugin: Could not find: libnccl-tuner.so. Using internal tuner plugin.
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO TUNER/Plugin: Could not find: libnccl-tuner.so. Using internal tuner plugin.
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO TUNER/Plugin: Could not find: libnccl-tuner.so. Using internal tuner plugin.
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO TUNER/Plugin: Could not find: libnccl-tuner.so. Using internal tuner plugin.
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO TUNER/Plugin: Could not find: libnccl-tuner.so. Using internal tuner plugin.
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO TUNER/Plugin: Failed to find ncclTunerPlugin_v4 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO TUNER/Plugin: Failed to find ncclTunerPlugin_v4 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO TUNER/Plugin: Using tuner plugin nccl_ofi_tuner
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO TUNER/Plugin: Using tuner plugin nccl_ofi_tuner
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO ncclCommInitRankConfig comm 0x55ad880448e0 rank 0 nranks 8 cudaDev 0 nvmlDev 0 busId 101c0 commId 0xc553500499762690 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO ncclCommInitRankConfig comm 0x5608cb6aab30 rank 3 nranks 8 cudaDev 3 nvmlDev 3 busId 201d0 commId 0xc553500499762690 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO TUNER/Plugin: Failed to find ncclTunerPlugin_v4 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83696:83984 [0] NCCL INFO Init timings - ncclCommInitRankConfig: rank 0 nranks 8 total 0.43 (kernels 0.19, alloc 0.10, bootstrap 0.00, allgathers 0.01, topo 0.05, graphs 0.00, connections 0.06, rest 0.02)
|
||
ywang29-p4d-debug-2-worker-0:83699:83982 [3] NCCL INFO Init timings - ncclCommInitRankConfig: rank 3 nranks 8 total 0.44 (kernels 0.20, alloc 0.10, bootstrap 0.00, allgathers 0.00, topo 0.05, graphs 0.00, connections 0.06, rest 0.02)
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO TUNER/Plugin: Using tuner plugin nccl_ofi_tuner
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO TUNER/Plugin: Failed to find ncclTunerPlugin_v4 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO TUNER/Plugin: Using tuner plugin nccl_ofi_tuner
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO ncclCommInitRankConfig comm 0x55e5169daea0 rank 2 nranks 8 cudaDev 2 nvmlDev 2 busId 201c0 commId 0xc553500499762690 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO ncclCommInitRankConfig comm 0x55ede9ca7de0 rank 1 nranks 8 cudaDev 1 nvmlDev 1 busId 101d0 commId 0xc553500499762690 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83698:83979 [2] NCCL INFO Init timings - ncclCommInitRankConfig: rank 2 nranks 8 total 0.46 (kernels 0.18, alloc 0.09, bootstrap 0.06, allgathers 0.00, topo 0.05, graphs 0.00, connections 0.06, rest 0.03)
|
||
ywang29-p4d-debug-2-worker-0:83697:83980 [1] NCCL INFO Init timings - ncclCommInitRankConfig: rank 1 nranks 8 total 0.46 (kernels 0.17, alloc 0.10, bootstrap 0.05, allgathers 0.01, topo 0.04, graphs 0.00, connections 0.06, rest 0.03)
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO TUNER/Plugin: Failed to find ncclTunerPlugin_v4 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO TUNER/Plugin: Using tuner plugin nccl_ofi_tuner
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO ncclCommInitRankConfig comm 0x55dcff54b080 rank 5 nranks 8 cudaDev 5 nvmlDev 5 busId 901d0 commId 0xc553500499762690 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO TUNER/Plugin: Failed to find ncclTunerPlugin_v4 symbol.
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO TUNER/Plugin: Using tuner plugin nccl_ofi_tuner
|
||
ywang29-p4d-debug-2-worker-0:83701:83981 [5] NCCL INFO Init timings - ncclCommInitRankConfig: rank 5 nranks 8 total 0.44 (kernels 0.20, alloc 0.10, bootstrap 0.00, allgathers 0.00, topo 0.05, graphs 0.00, connections 0.06, rest 0.02)
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO ncclCommInitRankConfig comm 0x55f5a30b2870 rank 7 nranks 8 cudaDev 7 nvmlDev 7 busId a01d0 commId 0xc553500499762690 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83703:83977 [7] NCCL INFO Init timings - ncclCommInitRankConfig: rank 7 nranks 8 total 0.68 (kernels 0.26, alloc 0.02, bootstrap 0.26, allgathers 0.01, topo 0.05, graphs 0.00, connections 0.06, rest 0.03)
|
||
[INFO|2026-04-30 16:19:36] llamafactory.hparams.parser:505 >> Process rank: 5, world size: 8, device: cuda:5, distributed training: True, compute dtype: torch.bfloat16
|
||
[INFO|2026-04-30 16:19:36] llamafactory.hparams.parser:505 >> Process rank: 1, world size: 8, device: cuda:1, distributed training: True, compute dtype: torch.bfloat16
|
||
[INFO|2026-04-30 16:19:36] llamafactory.hparams.parser:505 >> Process rank: 7, world size: 8, device: cuda:7, distributed training: True, compute dtype: torch.bfloat16
|
||
[INFO|2026-04-30 16:19:36] llamafactory.hparams.parser:505 >> Process rank: 2, world size: 8, device: cuda:2, distributed training: True, compute dtype: torch.bfloat16
|
||
[INFO|2026-04-30 16:19:36] llamafactory.hparams.parser:505 >> Process rank: 3, world size: 8, device: cuda:3, distributed training: True, compute dtype: torch.bfloat16
|
||
[INFO|2026-04-30 16:19:36] llamafactory.hparams.parser:505 >> Process rank: 6, world size: 8, device: cuda:6, distributed training: True, compute dtype: torch.bfloat16
|
||
[INFO|2026-04-30 16:19:36] llamafactory.hparams.parser:505 >> Process rank: 4, world size: 8, device: cuda:4, distributed training: True, compute dtype: torch.bfloat16
|
||
[INFO|2026-04-30 16:19:36] llamafactory.hparams.parser:505 >> Process rank: 0, world size: 8, device: cuda:0, distributed training: True, compute dtype: torch.bfloat16
|
||
[INFO|configuration_utils.py:670] 2026-04-30 16:19:36,260 >> loading configuration file config.json from cache at /root/.cache/huggingface/hub/models--Qwen--Qwen3-14B/snapshots/40c069824f4251a91eefaf281ebe4c544efd3e18/config.json
|
||
[INFO|configuration_utils.py:742] 2026-04-30 16:19:36,264 >> Model config Qwen3Config {
|
||
"architectures": [
|
||
"Qwen3ForCausalLM"
|
||
],
|
||
"attention_bias": false,
|
||
"attention_dropout": 0.0,
|
||
"bos_token_id": 151643,
|
||
"dtype": "bfloat16",
|
||
"eos_token_id": 151645,
|
||
"head_dim": 128,
|
||
"hidden_act": "silu",
|
||
"hidden_size": 5120,
|
||
"initializer_range": 0.02,
|
||
"intermediate_size": 17408,
|
||
"layer_types": [
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention"
|
||
],
|
||
"masked_layers": null,
|
||
"max_position_embeddings": 40960,
|
||
"max_window_layers": 40,
|
||
"model_type": "qwen3",
|
||
"num_attention_heads": 40,
|
||
"num_hidden_layers": 40,
|
||
"num_key_value_heads": 8,
|
||
"pad_token_id": null,
|
||
"rms_norm_eps": 1e-06,
|
||
"rope_parameters": {
|
||
"rope_theta": 1000000,
|
||
"rope_type": "default"
|
||
},
|
||
"sliding_window": null,
|
||
"sparsity_attn": null,
|
||
"sparsity_mlp": null,
|
||
"subnet_mode": null,
|
||
"subnet_type": null,
|
||
"threshold_attn": null,
|
||
"threshold_mlp": null,
|
||
"tie_word_embeddings": false,
|
||
"transformers_version": "5.2.0",
|
||
"use_cache": true,
|
||
"use_sliding_window": false,
|
||
"vocab_size": 151936
|
||
}
|
||
|
||
[INFO|configuration_utils.py:670] 2026-04-30 16:19:38,636 >> loading configuration file config.json from cache at /root/.cache/huggingface/hub/models--Qwen--Qwen3-14B/snapshots/40c069824f4251a91eefaf281ebe4c544efd3e18/config.json
|
||
[INFO|configuration_utils.py:742] 2026-04-30 16:19:38,637 >> Model config Qwen3Config {
|
||
"architectures": [
|
||
"Qwen3ForCausalLM"
|
||
],
|
||
"attention_bias": false,
|
||
"attention_dropout": 0.0,
|
||
"bos_token_id": 151643,
|
||
"dtype": "bfloat16",
|
||
"eos_token_id": 151645,
|
||
"head_dim": 128,
|
||
"hidden_act": "silu",
|
||
"hidden_size": 5120,
|
||
"initializer_range": 0.02,
|
||
"intermediate_size": 17408,
|
||
"layer_types": [
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention"
|
||
],
|
||
"masked_layers": null,
|
||
"max_position_embeddings": 40960,
|
||
"max_window_layers": 40,
|
||
"model_type": "qwen3",
|
||
"num_attention_heads": 40,
|
||
"num_hidden_layers": 40,
|
||
"num_key_value_heads": 8,
|
||
"pad_token_id": null,
|
||
"rms_norm_eps": 1e-06,
|
||
"rope_parameters": {
|
||
"rope_theta": 1000000,
|
||
"rope_type": "default"
|
||
},
|
||
"sliding_window": null,
|
||
"sparsity_attn": null,
|
||
"sparsity_mlp": null,
|
||
"subnet_mode": null,
|
||
"subnet_type": null,
|
||
"threshold_attn": null,
|
||
"threshold_mlp": null,
|
||
"tie_word_embeddings": false,
|
||
"transformers_version": "5.2.0",
|
||
"use_cache": true,
|
||
"use_sliding_window": false,
|
||
"vocab_size": 151936
|
||
}
|
||
|
||
[INFO|configuration_utils.py:670] 2026-04-30 16:19:38,753 >> loading configuration file config.json from cache at /root/.cache/huggingface/hub/models--Qwen--Qwen3-14B/snapshots/40c069824f4251a91eefaf281ebe4c544efd3e18/config.json
|
||
[INFO|configuration_utils.py:742] 2026-04-30 16:19:38,754 >> Model config Qwen3Config {
|
||
"architectures": [
|
||
"Qwen3ForCausalLM"
|
||
],
|
||
"attention_bias": false,
|
||
"attention_dropout": 0.0,
|
||
"bos_token_id": 151643,
|
||
"dtype": "bfloat16",
|
||
"eos_token_id": 151645,
|
||
"head_dim": 128,
|
||
"hidden_act": "silu",
|
||
"hidden_size": 5120,
|
||
"initializer_range": 0.02,
|
||
"intermediate_size": 17408,
|
||
"layer_types": [
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention"
|
||
],
|
||
"masked_layers": null,
|
||
"max_position_embeddings": 40960,
|
||
"max_window_layers": 40,
|
||
"model_type": "qwen3",
|
||
"num_attention_heads": 40,
|
||
"num_hidden_layers": 40,
|
||
"num_key_value_heads": 8,
|
||
"pad_token_id": null,
|
||
"rms_norm_eps": 1e-06,
|
||
"rope_parameters": {
|
||
"rope_theta": 1000000,
|
||
"rope_type": "default"
|
||
},
|
||
"sliding_window": null,
|
||
"sparsity_attn": null,
|
||
"sparsity_mlp": null,
|
||
"subnet_mode": null,
|
||
"subnet_type": null,
|
||
"threshold_attn": null,
|
||
"threshold_mlp": null,
|
||
"tie_word_embeddings": false,
|
||
"transformers_version": "5.2.0",
|
||
"use_cache": true,
|
||
"use_sliding_window": false,
|
||
"vocab_size": 151936
|
||
}
|
||
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 00/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 01/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 02/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 03/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 04/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 05/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 06/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 07/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 08/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 09/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 10/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 11/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 12/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 13/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 14/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 15/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 16/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 17/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 18/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 19/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 20/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 21/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 22/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Channel 23/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 00/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 01/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 02/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 03/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 04/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 05/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 06/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 07/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 08/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 09/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 10/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 11/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 12/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 13/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 14/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 15/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 16/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 17/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 18/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 19/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 20/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 21/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 22/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Channel 23/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 00/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 01/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 02/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 03/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 04/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 05/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 06/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 07/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 08/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 09/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 10/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 11/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 12/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 13/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 14/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 15/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 16/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 17/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 18/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 19/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 20/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 21/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 22/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Channel 23/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
[INFO|2026-04-30 16:19:40] llamafactory.data.loader:144 >> Loading dataset allenai/tulu-3-sft-personas-code...
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 00/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 01/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 02/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 03/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 04/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 05/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 06/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 07/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 08/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 09/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 10/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 11/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 12/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 13/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 14/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 15/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 16/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 17/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 18/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 19/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 20/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 21/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 22/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Channel 23/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 00/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 01/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 02/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 03/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 04/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 05/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 00/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 06/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 00/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 01/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 07/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 01/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 02/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 08/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 02/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 03/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 09/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 03/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 04/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 10/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 05/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 11/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 04/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 06/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 12/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 05/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 07/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 13/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 06/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 14/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 08/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 07/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 15/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 09/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 08/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 16/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 10/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 17/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 09/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 11/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 18/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 10/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 12/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 11/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 19/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 12/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 13/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 20/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 13/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 14/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 21/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 14/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 15/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 22/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 15/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Channel 23/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 16/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 16/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 17/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 17/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 18/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 18/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 19/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 19/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 20/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 20/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 21/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 21/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 22/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 22/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Channel 23/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Channel 23/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
[INFO|2026-04-30 16:19:41] llamafactory.data.loader:144 >> Sampled 10000 examples from dataset allenai/tulu-3-sft-personas-code.
|
||
[INFO|2026-04-30 16:19:41] llamafactory.data.loader:144 >> Loading dataset KKHYA/evol_codealpaca_converted...
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
[INFO|2026-04-30 16:19:42] llamafactory.data.loader:144 >> Sampled 10000 examples from dataset KKHYA/evol_codealpaca_converted.
|
||
[INFO|2026-04-30 16:19:42] llamafactory.data.loader:144 >> Loading dataset KKHYA/codefeedback_filtered_instructions_converted...
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
[INFO|2026-04-30 16:19:43] llamafactory.data.loader:144 >> Sampled 10000 examples from dataset KKHYA/codefeedback_filtered_instructions_converted.
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 00/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 01/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 02/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 03/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 04/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 05/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 06/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 07/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 08/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 09/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 10/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 11/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 12/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 13/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 14/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 15/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 16/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 17/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 18/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 19/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 20/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 21/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 22/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Channel 23/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84044 [6] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
ywang29-p4d-debug-2-worker-0:83698:84041 [2] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
ywang29-p4d-debug-2-worker-0:83697:84039 [1] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
ywang29-p4d-debug-2-worker-0:83696:84055 [0] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
ywang29-p4d-debug-2-worker-0:83703:84043 [7] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
ywang29-p4d-debug-2-worker-0:83701:84042 [5] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
ywang29-p4d-debug-2-worker-0:83700:84038 [4] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
ywang29-p4d-debug-2-worker-0:83699:84040 [3] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
training example:
|
||
input_ids:
|
||
[151644, 872, 198, 7985, 264, 10135, 729, 311, 11047, 279, 2790, 1372, 315, 8845, 16548, 553, 1674, 3767, 300, 4308, 304, 264, 2661, 3200, 504, 264, 1140, 315, 2432, 3059, 13, 8886, 2432, 1102, 374, 15251, 438, 264, 10997, 448, 6894, 330, 5117, 26532, 497, 330, 13757, 26532, 497, 330, 5117, 96244, 497, 323, 330, 13757, 96244, 3263, 1674, 3767, 300, 4308, 1410, 387, 2987, 279, 2114, 476, 3123, 2083, 304, 894, 2432, 13, 576, 1946, 374, 264, 1140, 315, 1741, 2432, 1102, 57514, 11, 323, 279, 2550, 1265, 387, 458, 7546, 14064, 279, 2790, 1372, 315, 8845, 16548, 553, 1674, 3767, 300, 4308, 382, 2505, 510, 12, 1565, 6347, 13576, 44622, 362, 1140, 315, 57514, 11, 1380, 1817, 10997, 5610, 510, 220, 481, 330, 5117, 26532, 1, 320, 917, 1648, 576, 829, 315, 279, 2114, 2083, 624, 220, 481, 330, 13757, 26532, 1, 320, 917, 1648, 576, 829, 315, 279, 3123, 2083, 624, 220, 481, 330, 5117, 96244, 1, 320, 396, 1648, 576, 1372, 315, 8845, 16548, 553, 279, 2114, 2083, 624, 220, 481, 330, 13757, 96244, 1, 320, 396, 1648, 576, 1372, 315, 8845, 16548, 553, 279, 3123, 2083, 382, 5097, 510, 12, 1527, 7546, 14064, 279, 2790, 1372, 315, 8845, 16548, 553, 1674, 3767, 300, 4308, 382, 13314, 510, 73594, 12669, 198, 6347, 13576, 284, 2278, 262, 5212, 5117, 26532, 788, 330, 2101, 3767, 300, 4308, 497, 330, 13757, 26532, 788, 330, 14597, 32, 497, 330, 5117, 96244, 788, 220, 17, 11, 330, 13757, 96244, 788, 220, 16, 1583, 262, 5212, 5117, 26532, 788, 330, 14597, 33, 497, 330, 13757, 26532, 788, 330, 2101, 3767, 300, 4308, 497, 330, 5117, 96244, 788, 220, 18, 11, 330, 13757, 96244, 788, 220, 17, 1583, 262, 5212, 5117, 26532, 788, 330, 2101, 3767, 300, 4308, 497, 330, 13757, 26532, 788, 330, 14597, 34, 497, 330, 5117, 96244, 788, 220, 16, 11, 330, 13757, 96244, 788, 220, 16, 1583, 262, 5212, 5117, 26532, 788, 330, 14597, 35, 497, 330, 13757, 26532, 788, 330, 2101, 3767, 300, 4308, 497, 330, 5117, 96244, 788, 220, 15, 11, 330, 13757, 96244, 788, 220, 18, 532, 921, 2, 31021, 9258, 25, 220, 23, 198, 73594, 151645, 198, 151644, 77091, 198, 750, 11047, 8418, 3767, 300, 4308, 96244, 25401, 13576, 982, 262, 2790, 96244, 284, 220, 15, 198, 262, 369, 2432, 304, 2432, 13576, 510, 286, 421, 2432, 1183, 5117, 26532, 1341, 621, 330, 2101, 3767, 300, 4308, 4660, 310, 2790, 96244, 1421, 2432, 1183, 5117, 96244, 7026, 286, 4409, 2432, 1183, 13757, 26532, 1341, 621, 330, 2101, 3767, 300, 4308, 4660, 310, 2790, 96244, 1421, 2432, 1183, 13757, 96244, 7026, 262, 470, 2790, 96244, 151645, 198]
|
||
inputs:
|
||
<|im_start|>user
|
||
Write a python function to calculate the total number of goals scored by Alanyaspor in a given season from a list of match results. Each match result is represented as a dictionary with keys "home_team", "away_team", "home_goals", and "away_goals". Alanyaspor could be either the home or away team in any match. The input is a list of such match result dictionaries, and the output should be an integer representing the total number of goals scored by Alanyaspor.
|
||
|
||
Input:
|
||
- `match_results`: A list of dictionaries, where each dictionary contains:
|
||
- "home_team" (string): The name of the home team.
|
||
- "away_team" (string): The name of the away team.
|
||
- "home_goals" (int): The number of goals scored by the home team.
|
||
- "away_goals" (int): The number of goals scored by the away team.
|
||
|
||
Output:
|
||
- An integer representing the total number of goals scored by Alanyaspor.
|
||
|
||
Example:
|
||
```python
|
||
match_results = [
|
||
{"home_team": "Alanyaspor", "away_team": "TeamA", "home_goals": 2, "away_goals": 1},
|
||
{"home_team": "TeamB", "away_team": "Alanyaspor", "home_goals": 3, "away_goals": 2},
|
||
{"home_team": "Alanyaspor", "away_team": "TeamC", "home_goals": 1, "away_goals": 1},
|
||
{"home_team": "TeamD", "away_team": "Alanyaspor", "home_goals": 0, "away_goals": 3}
|
||
]
|
||
# Expected Output: 8
|
||
```<|im_end|>
|
||
<|im_start|>assistant
|
||
def calculate_alanyaspor_goals(match_results):
|
||
total_goals = 0
|
||
for match in match_results:
|
||
if match["home_team"] == "Alanyaspor":
|
||
total_goals += match["home_goals"]
|
||
elif match["away_team"] == "Alanyaspor":
|
||
total_goals += match["away_goals"]
|
||
return total_goals<|im_end|>
|
||
|
||
label_ids:
|
||
[-100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, -100, 750, 11047, 8418, 3767, 300, 4308, 96244, 25401, 13576, 982, 262, 2790, 96244, 284, 220, 15, 198, 262, 369, 2432, 304, 2432, 13576, 510, 286, 421, 2432, 1183, 5117, 26532, 1341, 621, 330, 2101, 3767, 300, 4308, 4660, 310, 2790, 96244, 1421, 2432, 1183, 5117, 96244, 7026, 286, 4409, 2432, 1183, 13757, 26532, 1341, 621, 330, 2101, 3767, 300, 4308, 4660, 310, 2790, 96244, 1421, 2432, 1183, 13757, 96244, 7026, 262, 470, 2790, 96244, 151645, 198]
|
||
labels:
|
||
def calculate_alanyaspor_goals(match_results):
|
||
total_goals = 0
|
||
for match in match_results:
|
||
if match["home_team"] == "Alanyaspor":
|
||
total_goals += match["home_goals"]
|
||
elif match["away_team"] == "Alanyaspor":
|
||
total_goals += match["away_goals"]
|
||
return total_goals<|im_end|>
|
||
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
Repo card metadata block was not found. Setting CardData to empty.
|
||
[INFO|configuration_utils.py:670] 2026-04-30 16:19:47,605 >> loading configuration file config.json from cache at /root/.cache/huggingface/hub/models--Qwen--Qwen3-14B/snapshots/40c069824f4251a91eefaf281ebe4c544efd3e18/config.json
|
||
[INFO|configuration_utils.py:742] 2026-04-30 16:19:47,605 >> Model config Qwen3Config {
|
||
"architectures": [
|
||
"Qwen3ForCausalLM"
|
||
],
|
||
"attention_bias": false,
|
||
"attention_dropout": 0.0,
|
||
"bos_token_id": 151643,
|
||
"dtype": "bfloat16",
|
||
"eos_token_id": 151645,
|
||
"head_dim": 128,
|
||
"hidden_act": "silu",
|
||
"hidden_size": 5120,
|
||
"initializer_range": 0.02,
|
||
"intermediate_size": 17408,
|
||
"layer_types": [
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention",
|
||
"full_attention"
|
||
],
|
||
"masked_layers": null,
|
||
"max_position_embeddings": 40960,
|
||
"max_window_layers": 40,
|
||
"model_type": "qwen3",
|
||
"num_attention_heads": 40,
|
||
"num_hidden_layers": 40,
|
||
"num_key_value_heads": 8,
|
||
"pad_token_id": null,
|
||
"rms_norm_eps": 1e-06,
|
||
"rope_parameters": {
|
||
"rope_theta": 1000000,
|
||
"rope_type": "default"
|
||
},
|
||
"sliding_window": null,
|
||
"sparsity_attn": null,
|
||
"sparsity_mlp": null,
|
||
"subnet_mode": null,
|
||
"subnet_type": null,
|
||
"threshold_attn": null,
|
||
"threshold_mlp": null,
|
||
"tie_word_embeddings": false,
|
||
"transformers_version": "5.2.0",
|
||
"use_cache": true,
|
||
"use_sliding_window": false,
|
||
"vocab_size": 151936
|
||
}
|
||
|
||
[INFO|2026-04-30 16:19:47] llamafactory.model.model_utils.kv_cache:144 >> KV cache is disabled during training.
|
||
[INFO|modeling_utils.py:710] 2026-04-30 16:19:48,179 >> loading weights file model.safetensors from cache at /root/.cache/huggingface/hub/models--Qwen--Qwen3-14B/snapshots/40c069824f4251a91eefaf281ebe4c544efd3e18/model.safetensors.index.json
|
||
[INFO|modeling_utils.py:779] 2026-04-30 16:19:48,179 >> Will use dtype=torch.bfloat16 as defined in model's config object
|
||
[INFO|modeling_utils.py:3560] 2026-04-30 16:19:48,179 >> Detected DeepSpeed ZeRO-3: activating zero.init() for this model
|
||
[INFO|configuration_utils.py:1014] 2026-04-30 16:19:48,191 >> Generate config GenerationConfig {
|
||
"bos_token_id": 151643,
|
||
"eos_token_id": 151645,
|
||
"output_attentions": false,
|
||
"output_hidden_states": false,
|
||
"use_cache": false
|
||
}
|
||
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to lm_head.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to lm_head.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to lm_head.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to lm_head.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to lm_head.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to lm_head.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to lm_head.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tokenizer has new PAD/BOS/EOS tokens that differ from the model config and generation config. The model config and generation config were aligned accordingly, being updated with the tokenizer's values. Updated tokens: {'bos_token_id': None, 'pad_token_id': 151643}.
|
||
The tokenizer has new PAD/BOS/EOS tokens that differ from the model config and generation config. The model config and generation config were aligned accordingly, being updated with the tokenizer's values. Updated tokens: {'bos_token_id': None, 'pad_token_id': 151643}.
|
||
The tokenizer has new PAD/BOS/EOS tokens that differ from the model config and generation config. The model config and generation config were aligned accordingly, being updated with the tokenizer's values. Updated tokens: {'bos_token_id': None, 'pad_token_id': 151643}.
|
||
The tokenizer has new PAD/BOS/EOS tokens that differ from the model config and generation config. The model config and generation config were aligned accordingly, being updated with the tokenizer's values. Updated tokens: {'bos_token_id': None, 'pad_token_id': 151643}.
|
||
The tokenizer has new PAD/BOS/EOS tokens that differ from the model config and generation config. The model config and generation config were aligned accordingly, being updated with the tokenizer's values. Updated tokens: {'bos_token_id': None, 'pad_token_id': 151643}.
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,952 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.0.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.1.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.2.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,953 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.3.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.4.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.5.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,954 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.6.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.7.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.8.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.9.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,955 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.10.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.11.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.12.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.13.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,956 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.14.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.15.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.16.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,957 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.17.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.18.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.19.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.20.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,958 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.21.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.22.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.23.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.24.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,959 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.25.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.26.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.27.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,960 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.28.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.29.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.30.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.31.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,961 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.32.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.33.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.34.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,962 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.35.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.36.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.37.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.38.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.v_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,963 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.o_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,964 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.q_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,964 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.self_attn.k_norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,964 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.gate_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,964 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.up_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,964 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.mlp.down_proj.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,964 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.input_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,964 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.layers.39.post_attention_layernorm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,964 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to model.norm.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
[WARNING|modeling_utils.py:2496] 2026-04-30 16:19:58,964 >> The tied weights mapping and config for this model specifies to tie model.embed_tokens.weight to lm_head.weight, but both are present in the checkpoints, so we will NOT tie them. You should update the config with `tie_word_embeddings=False` to silence this warning
|
||
The tokenizer has new PAD/BOS/EOS tokens that differ from the model config and generation config. The model config and generation config were aligned accordingly, being updated with the tokenizer's values. Updated tokens: {'bos_token_id': None, 'pad_token_id': 151643}.
|
||
The tokenizer has new PAD/BOS/EOS tokens that differ from the model config and generation config. The model config and generation config were aligned accordingly, being updated with the tokenizer's values. Updated tokens: {'bos_token_id': None, 'pad_token_id': 151643}.
|
||
[INFO|configuration_utils.py:967] 2026-04-30 16:19:59,095 >> loading configuration file generation_config.json from cache at /root/.cache/huggingface/hub/models--Qwen--Qwen3-14B/snapshots/40c069824f4251a91eefaf281ebe4c544efd3e18/generation_config.json
|
||
[INFO|configuration_utils.py:1014] 2026-04-30 16:19:59,096 >> Generate config GenerationConfig {
|
||
"bos_token_id": 151643,
|
||
"do_sample": true,
|
||
"eos_token_id": [
|
||
151645,
|
||
151643
|
||
],
|
||
"pad_token_id": 151643,
|
||
"temperature": 0.6,
|
||
"top_k": 20,
|
||
"top_p": 0.95
|
||
}
|
||
|
||
ywang29-p4d-debug-2-worker-0:83699:83699 [3] NCCL INFO Comm config Blocking set to 1
|
||
ywang29-p4d-debug-2-worker-0:83697:83697 [1] NCCL INFO Comm config Blocking set to 1
|
||
[INFO|dynamic_module_utils.py:406] 2026-04-30 16:19:59,198 >> Could not locate the custom_generate/generate.py inside Qwen/Qwen3-14B.
|
||
[INFO|2026-04-30 16:19:59] llamafactory.model.model_utils.checkpointing:144 >> Gradient checkpointing enabled.
|
||
[INFO|2026-04-30 16:19:59] llamafactory.model.model_utils.attention:144 >> Using torch SDPA for faster training and inference.
|
||
[INFO|2026-04-30 16:19:59] llamafactory.model.adapter:144 >> DeepSpeed ZeRO3 detected, remaining trainable params in float32.
|
||
[INFO|2026-04-30 16:19:59] llamafactory.model.adapter:144 >> Fine-tuning method: Full
|
||
[INFO|2026-04-30 16:19:59] llamafactory.model.loader:144 >> trainable params: 14,768,307,200 || all params: 14,768,307,200 || trainable%: 100.0000
|
||
ywang29-p4d-debug-2-worker-0:83701:83701 [5] NCCL INFO Comm config Blocking set to 1
|
||
ywang29-p4d-debug-2-worker-0:83700:83700 [4] NCCL INFO Comm config Blocking set to 1
|
||
ywang29-p4d-debug-2-worker-0:83702:83702 [6] NCCL INFO Comm config Blocking set to 1
|
||
[WARNING|2026-04-30 16:19:59] llamafactory.train.callbacks:155 >> Previous trainer log in this folder will be deleted.
|
||
[WARNING|trainer_utils.py:1234] 2026-04-30 16:19:59,381 >> The tokenizer has new PAD/BOS/EOS tokens that differ from the model config and generation config. The model config and generation config were aligned accordingly, being updated with the tokenizer's values. Updated tokens: {'bos_token_id': None, 'pad_token_id': 151643}.
|
||
ywang29-p4d-debug-2-worker-0:83703:83703 [7] NCCL INFO Comm config Blocking set to 1
|
||
ywang29-p4d-debug-2-worker-0:83698:83698 [2] NCCL INFO Comm config Blocking set to 1
|
||
ywang29-p4d-debug-2-worker-0:83696:83696 [0] NCCL INFO Comm config Blocking set to 1
|
||
ywang29-p4d-debug-2-worker-0:83699:84128 [3] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83697:84131 [1] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83701:84134 [5] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Using network Socket
|
||
ywang29-p4d-debug-2-worker-0:83697:84131 [1] NCCL INFO Using network Socket
|
||
ywang29-p4d-debug-2-worker-0:83701:84134 [5] NCCL INFO Using network Socket
|
||
ywang29-p4d-debug-2-worker-0:83699:84128 [3] NCCL INFO Using network Socket
|
||
ywang29-p4d-debug-2-worker-0:83702:84140 [6] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83698:84146 [2] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83703:84143 [7] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83702:84140 [6] NCCL INFO Using network Socket
|
||
ywang29-p4d-debug-2-worker-0:83698:84146 [2] NCCL INFO Using network Socket
|
||
ywang29-p4d-debug-2-worker-0:83703:84143 [7] NCCL INFO Using network Socket
|
||
ywang29-p4d-debug-2-worker-0:83700:84137 [4] NCCL INFO Assigned NET plugin Socket to comm
|
||
ywang29-p4d-debug-2-worker-0:83700:84137 [4] NCCL INFO Using network Socket
|
||
ywang29-p4d-debug-2-worker-0:83701:84134 [5] NCCL INFO ncclCommSplit comm 0x55dd0ac0a840 rank 5 nranks 8 cudaDev 5 nvmlDev 5 busId 901d0 parent 0x55dcff54b080 splitCount 1 color 1266629538 key 5- Init START
|
||
ywang29-p4d-debug-2-worker-0:83699:84128 [3] NCCL INFO ncclCommSplit comm 0x5608d6d641f0 rank 3 nranks 8 cudaDev 3 nvmlDev 3 busId 201d0 parent 0x5608cb6aab30 splitCount 1 color 1266629538 key 3- Init START
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO ncclCommSplit comm 0x55ad937092f0 rank 0 nranks 8 cudaDev 0 nvmlDev 0 busId 101c0 parent 0x55ad880448e0 splitCount 1 color 1266629538 key 0- Init START
|
||
ywang29-p4d-debug-2-worker-0:83697:84131 [1] NCCL INFO ncclCommSplit comm 0x55edf53941b0 rank 1 nranks 8 cudaDev 1 nvmlDev 1 busId 101d0 parent 0x55ede9ca7de0 splitCount 1 color 1266629538 key 1- Init START
|
||
ywang29-p4d-debug-2-worker-0:83700:84137 [4] NCCL INFO ncclCommSplit comm 0x562ca54bbc00 rank 4 nranks 8 cudaDev 4 nvmlDev 4 busId 901c0 parent 0x562c99db2980 splitCount 1 color 1266629538 key 4- Init START
|
||
ywang29-p4d-debug-2-worker-0:83698:84146 [2] NCCL INFO ncclCommSplit comm 0x55e5220b9800 rank 2 nranks 8 cudaDev 2 nvmlDev 2 busId 201c0 parent 0x55e5169daea0 splitCount 1 color 1266629538 key 2- Init START
|
||
ywang29-p4d-debug-2-worker-0:83702:84140 [6] NCCL INFO ncclCommSplit comm 0x55a9dbb4c8b0 rank 6 nranks 8 cudaDev 6 nvmlDev 6 busId a01c0 parent 0x55a9d0483a80 splitCount 1 color 1266629538 key 6- Init START
|
||
ywang29-p4d-debug-2-worker-0:83703:84143 [7] NCCL INFO ncclCommSplit comm 0x55f5ae797530 rank 7 nranks 8 cudaDev 7 nvmlDev 7 busId a01d0 parent 0x55f5a30b2870 splitCount 1 color 1266629538 key 7- Init START
|
||
ywang29-p4d-debug-2-worker-0:83697:84131 [1] NCCL INFO Setting affinity for GPU 1 to 0-23,48-71
|
||
ywang29-p4d-debug-2-worker-0:83703:84143 [7] NCCL INFO Setting affinity for GPU 7 to 24-47,72-95
|
||
ywang29-p4d-debug-2-worker-0:83697:84131 [1] NCCL INFO NVLS multicast support is not available on dev 1 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Setting affinity for GPU 0 to 0-23,48-71
|
||
ywang29-p4d-debug-2-worker-0:83703:84143 [7] NCCL INFO NVLS multicast support is not available on dev 7 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO NVLS multicast support is not available on dev 0 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83702:84140 [6] NCCL INFO Setting affinity for GPU 6 to 24-47,72-95
|
||
ywang29-p4d-debug-2-worker-0:83702:84140 [6] NCCL INFO NVLS multicast support is not available on dev 6 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83698:84146 [2] NCCL INFO Setting affinity for GPU 2 to 0-23,48-71
|
||
ywang29-p4d-debug-2-worker-0:83698:84146 [2] NCCL INFO NVLS multicast support is not available on dev 2 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83700:84137 [4] NCCL INFO Setting affinity for GPU 4 to 24-47,72-95
|
||
ywang29-p4d-debug-2-worker-0:83701:84134 [5] NCCL INFO Setting affinity for GPU 5 to 24-47,72-95
|
||
ywang29-p4d-debug-2-worker-0:83701:84134 [5] NCCL INFO NVLS multicast support is not available on dev 5 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83700:84137 [4] NCCL INFO NVLS multicast support is not available on dev 4 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83699:84128 [3] NCCL INFO Setting affinity for GPU 3 to 0-23,48-71
|
||
ywang29-p4d-debug-2-worker-0:83699:84128 [3] NCCL INFO NVLS multicast support is not available on dev 3 (NVLS_NCHANNELS 0)
|
||
ywang29-p4d-debug-2-worker-0:83699:84128 [3] NCCL INFO comm 0x5608d6d641f0 rank 3 nRanks 8 nNodes 1 localRanks 8 localRank 3 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83698:84146 [2] NCCL INFO comm 0x55e5220b9800 rank 2 nRanks 8 nNodes 1 localRanks 8 localRank 2 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83701:84134 [5] NCCL INFO comm 0x55dd0ac0a840 rank 5 nRanks 8 nNodes 1 localRanks 8 localRank 5 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83697:84131 [1] NCCL INFO comm 0x55edf53941b0 rank 1 nRanks 8 nNodes 1 localRanks 8 localRank 1 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO comm 0x55ad937092f0 rank 0 nRanks 8 nNodes 1 localRanks 8 localRank 0 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83703:84143 [7] NCCL INFO comm 0x55f5ae797530 rank 7 nRanks 8 nNodes 1 localRanks 8 localRank 7 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83700:84137 [4] NCCL INFO comm 0x562ca54bbc00 rank 4 nRanks 8 nNodes 1 localRanks 8 localRank 4 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83702:84140 [6] NCCL INFO comm 0x55a9dbb4c8b0 rank 6 nRanks 8 nNodes 1 localRanks 8 localRank 6 MNNVL 0
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 00/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 01/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 02/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 03/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 04/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 05/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83698:84146 [2] NCCL INFO Trees [0] 3/-1/-1->2->1 [1] 3/-1/-1->2->1 [2] 3/-1/-1->2->1 [3] 3/-1/-1->2->1 [4] 3/-1/-1->2->1 [5] 3/-1/-1->2->1 [6] 3/-1/-1->2->1 [7] 3/-1/-1->2->1 [8] 3/-1/-1->2->1 [9] 3/-1/-1->2->1 [10] 3/-1/-1->2->1 [11] 3/-1/-1->2->1 [12] 3/-1/-1->2->1 [13] 3/-1/-1->2->1 [14] 3/-1/-1->2->1 [15] 3/-1/-1->2->1 [16] 3/-1/-1->2->1 [17] 3/-1/-1->2->1 [18] 3/-1/-1->2->1 [19] 3/-1/-1->2->1 [20] 3/-1/-1->2->1 [21] 3/-1/-1->2->1 [22] 3/-1/-1->2->1 [23] 3/-1/-1->2->1
|
||
ywang29-p4d-debug-2-worker-0:83699:84128 [3] NCCL INFO Trees [0] 4/-1/-1->3->2 [1] 4/-1/-1->3->2 [2] 4/-1/-1->3->2 [3] 4/-1/-1->3->2 [4] 4/-1/-1->3->2 [5] 4/-1/-1->3->2 [6] 4/-1/-1->3->2 [7] 4/-1/-1->3->2 [8] 4/-1/-1->3->2 [9] 4/-1/-1->3->2 [10] 4/-1/-1->3->2 [11] 4/-1/-1->3->2 [12] 4/-1/-1->3->2 [13] 4/-1/-1->3->2 [14] 4/-1/-1->3->2 [15] 4/-1/-1->3->2 [16] 4/-1/-1->3->2 [17] 4/-1/-1->3->2 [18] 4/-1/-1->3->2 [19] 4/-1/-1->3->2 [20] 4/-1/-1->3->2 [21] 4/-1/-1->3->2 [22] 4/-1/-1->3->2 [23] 4/-1/-1->3->2
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 06/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83698:84146 [2] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83701:84134 [5] NCCL INFO Trees [0] 6/-1/-1->5->4 [1] 6/-1/-1->5->4 [2] 6/-1/-1->5->4 [3] 6/-1/-1->5->4 [4] 6/-1/-1->5->4 [5] 6/-1/-1->5->4 [6] 6/-1/-1->5->4 [7] 6/-1/-1->5->4 [8] 6/-1/-1->5->4 [9] 6/-1/-1->5->4 [10] 6/-1/-1->5->4 [11] 6/-1/-1->5->4 [12] 6/-1/-1->5->4 [13] 6/-1/-1->5->4 [14] 6/-1/-1->5->4 [15] 6/-1/-1->5->4 [16] 6/-1/-1->5->4 [17] 6/-1/-1->5->4 [18] 6/-1/-1->5->4 [19] 6/-1/-1->5->4 [20] 6/-1/-1->5->4 [21] 6/-1/-1->5->4 [22] 6/-1/-1->5->4 [23] 6/-1/-1->5->4
|
||
ywang29-p4d-debug-2-worker-0:83699:84128 [3] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83697:84131 [1] NCCL INFO Trees [0] 2/-1/-1->1->0 [1] 2/-1/-1->1->0 [2] 2/-1/-1->1->0 [3] 2/-1/-1->1->0 [4] 2/-1/-1->1->0 [5] 2/-1/-1->1->0 [6] 2/-1/-1->1->0 [7] 2/-1/-1->1->0 [8] 2/-1/-1->1->0 [9] 2/-1/-1->1->0 [10] 2/-1/-1->1->0 [11] 2/-1/-1->1->0 [12] 2/-1/-1->1->0 [13] 2/-1/-1->1->0 [14] 2/-1/-1->1->0 [15] 2/-1/-1->1->0 [16] 2/-1/-1->1->0 [17] 2/-1/-1->1->0 [18] 2/-1/-1->1->0 [19] 2/-1/-1->1->0 [20] 2/-1/-1->1->0 [21] 2/-1/-1->1->0 [22] 2/-1/-1->1->0 [23] 2/-1/-1->1->0
|
||
ywang29-p4d-debug-2-worker-0:83701:84134 [5] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 07/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83697:84131 [1] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83703:84143 [7] NCCL INFO Trees [0] -1/-1/-1->7->6 [1] -1/-1/-1->7->6 [2] -1/-1/-1->7->6 [3] -1/-1/-1->7->6 [4] -1/-1/-1->7->6 [5] -1/-1/-1->7->6 [6] -1/-1/-1->7->6 [7] -1/-1/-1->7->6 [8] -1/-1/-1->7->6 [9] -1/-1/-1->7->6 [10] -1/-1/-1->7->6 [11] -1/-1/-1->7->6 [12] -1/-1/-1->7->6 [13] -1/-1/-1->7->6 [14] -1/-1/-1->7->6 [15] -1/-1/-1->7->6 [16] -1/-1/-1->7->6 [17] -1/-1/-1->7->6 [18] -1/-1/-1->7->6 [19] -1/-1/-1->7->6 [20] -1/-1/-1->7->6 [21] -1/-1/-1->7->6 [22] -1/-1/-1->7->6 [23] -1/-1/-1->7->6
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 08/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83703:84143 [7] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 09/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 10/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 11/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 12/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 13/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 14/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83700:84137 [4] NCCL INFO Trees [0] 5/-1/-1->4->3 [1] 5/-1/-1->4->3 [2] 5/-1/-1->4->3 [3] 5/-1/-1->4->3 [4] 5/-1/-1->4->3 [5] 5/-1/-1->4->3 [6] 5/-1/-1->4->3 [7] 5/-1/-1->4->3 [8] 5/-1/-1->4->3 [9] 5/-1/-1->4->3 [10] 5/-1/-1->4->3 [11] 5/-1/-1->4->3 [12] 5/-1/-1->4->3 [13] 5/-1/-1->4->3 [14] 5/-1/-1->4->3 [15] 5/-1/-1->4->3 [16] 5/-1/-1->4->3 [17] 5/-1/-1->4->3 [18] 5/-1/-1->4->3 [19] 5/-1/-1->4->3 [20] 5/-1/-1->4->3 [21] 5/-1/-1->4->3 [22] 5/-1/-1->4->3 [23] 5/-1/-1->4->3
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 15/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83700:84137 [4] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83702:84140 [6] NCCL INFO Trees [0] 7/-1/-1->6->5 [1] 7/-1/-1->6->5 [2] 7/-1/-1->6->5 [3] 7/-1/-1->6->5 [4] 7/-1/-1->6->5 [5] 7/-1/-1->6->5 [6] 7/-1/-1->6->5 [7] 7/-1/-1->6->5 [8] 7/-1/-1->6->5 [9] 7/-1/-1->6->5 [10] 7/-1/-1->6->5 [11] 7/-1/-1->6->5 [12] 7/-1/-1->6->5 [13] 7/-1/-1->6->5 [14] 7/-1/-1->6->5 [15] 7/-1/-1->6->5 [16] 7/-1/-1->6->5 [17] 7/-1/-1->6->5 [18] 7/-1/-1->6->5 [19] 7/-1/-1->6->5 [20] 7/-1/-1->6->5 [21] 7/-1/-1->6->5 [22] 7/-1/-1->6->5 [23] 7/-1/-1->6->5
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 16/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83702:84140 [6] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 17/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 18/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 19/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 20/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 21/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 22/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Channel 23/24 : 0 1 2 3 4 5 6 7
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Trees [0] 1/-1/-1->0->-1 [1] 1/-1/-1->0->-1 [2] 1/-1/-1->0->-1 [3] 1/-1/-1->0->-1 [4] 1/-1/-1->0->-1 [5] 1/-1/-1->0->-1 [6] 1/-1/-1->0->-1 [7] 1/-1/-1->0->-1 [8] 1/-1/-1->0->-1 [9] 1/-1/-1->0->-1 [10] 1/-1/-1->0->-1 [11] 1/-1/-1->0->-1 [12] 1/-1/-1->0->-1 [13] 1/-1/-1->0->-1 [14] 1/-1/-1->0->-1 [15] 1/-1/-1->0->-1 [16] 1/-1/-1->0->-1 [17] 1/-1/-1->0->-1 [18] 1/-1/-1->0->-1 [19] 1/-1/-1->0->-1 [20] 1/-1/-1->0->-1 [21] 1/-1/-1->0->-1 [22] 1/-1/-1->0->-1 [23] 1/-1/-1->0->-1
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO P2P Chunksize set to 524288
|
||
ywang29-p4d-debug-2-worker-0:83702:84150 [6] NCCL INFO [Proxy Service] Device 6 CPU core 78
|
||
ywang29-p4d-debug-2-worker-0:83702:84151 [6] NCCL INFO [Proxy Service UDS] Device 6 CPU core 81
|
||
ywang29-p4d-debug-2-worker-0:83703:84153 [7] NCCL INFO [Proxy Service UDS] Device 7 CPU core 83
|
||
ywang29-p4d-debug-2-worker-0:83703:84152 [7] NCCL INFO [Proxy Service] Device 7 CPU core 82
|
||
ywang29-p4d-debug-2-worker-0:83698:84155 [2] NCCL INFO [Proxy Service UDS] Device 2 CPU core 55
|
||
ywang29-p4d-debug-2-worker-0:83698:84154 [2] NCCL INFO [Proxy Service] Device 2 CPU core 54
|
||
ywang29-p4d-debug-2-worker-0:83700:84156 [4] NCCL INFO [Proxy Service] Device 4 CPU core 76
|
||
ywang29-p4d-debug-2-worker-0:83700:84157 [4] NCCL INFO [Proxy Service UDS] Device 4 CPU core 84
|
||
ywang29-p4d-debug-2-worker-0:83701:84158 [5] NCCL INFO [Proxy Service] Device 5 CPU core 37
|
||
ywang29-p4d-debug-2-worker-0:83701:84159 [5] NCCL INFO [Proxy Service UDS] Device 5 CPU core 88
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Check P2P Type isAllDirectP2p 1 directMode 0
|
||
ywang29-p4d-debug-2-worker-0:83696:84161 [0] NCCL INFO [Proxy Service UDS] Device 0 CPU core 64
|
||
ywang29-p4d-debug-2-worker-0:83696:84160 [0] NCCL INFO [Proxy Service] Device 0 CPU core 13
|
||
ywang29-p4d-debug-2-worker-0:83697:84162 [1] NCCL INFO [Proxy Service] Device 1 CPU core 20
|
||
ywang29-p4d-debug-2-worker-0:83697:84163 [1] NCCL INFO [Proxy Service UDS] Device 1 CPU core 21
|
||
ywang29-p4d-debug-2-worker-0:83699:84164 [3] NCCL INFO [Proxy Service] Device 3 CPU core 9
|
||
ywang29-p4d-debug-2-worker-0:83699:84165 [3] NCCL INFO [Proxy Service UDS] Device 3 CPU core 11
|
||
ywang29-p4d-debug-2-worker-0:83698:84146 [2] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83698:84146 [2] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83700:84137 [4] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83700:84137 [4] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83702:84140 [6] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83702:84140 [6] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83699:84128 [3] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83699:84128 [3] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83697:84131 [1] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83697:84131 [1] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83703:84143 [7] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83703:84143 [7] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO CC Off, workFifoBytes 1048576
|
||
ywang29-p4d-debug-2-worker-0:83701:84134 [5] NCCL INFO threadThresholds 8/8/64 | 64/8/64 | 512 | 512
|
||
ywang29-p4d-debug-2-worker-0:83701:84134 [5] NCCL INFO 24 coll channels, 24 collnet channels, 0 nvls channels, 32 p2p channels, 32 p2p channels per peer
|
||
ywang29-p4d-debug-2-worker-0:83703:84143 [7] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83699:84128 [3] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83703:84143 [7] NCCL INFO ncclCommSplit comm 0x55f5ae797530 rank 7 nranks 8 cudaDev 7 nvmlDev 7 busId a01d0 parent 0x55f5a30b2870 splitCount 1 color 1266629538 key 7 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83699:84128 [3] NCCL INFO ncclCommSplit comm 0x5608d6d641f0 rank 3 nranks 8 cudaDev 3 nvmlDev 3 busId 201d0 parent 0x5608cb6aab30 splitCount 1 color 1266629538 key 3 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83697:84131 [1] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83698:84146 [2] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83703:84143 [7] NCCL INFO Init timings - ncclCommSplit: rank 7 nranks 8 total 0.45 (kernels 0.00, alloc 0.00, bootstrap 0.00, allgathers 0.00, topo 0.03, graphs 0.00, connections 0.07, rest 0.35)
|
||
ywang29-p4d-debug-2-worker-0:83697:84131 [1] NCCL INFO ncclCommSplit comm 0x55edf53941b0 rank 1 nranks 8 cudaDev 1 nvmlDev 1 busId 101d0 parent 0x55ede9ca7de0 splitCount 1 color 1266629538 key 1 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83698:84146 [2] NCCL INFO ncclCommSplit comm 0x55e5220b9800 rank 2 nranks 8 cudaDev 2 nvmlDev 2 busId 201c0 parent 0x55e5169daea0 splitCount 1 color 1266629538 key 2 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83699:84128 [3] NCCL INFO Init timings - ncclCommSplit: rank 3 nranks 8 total 0.78 (kernels 0.00, alloc 0.00, bootstrap 0.00, allgathers 0.00, topo 0.03, graphs 0.00, connections 0.06, rest 0.69)
|
||
ywang29-p4d-debug-2-worker-0:83697:84131 [1] NCCL INFO Init timings - ncclCommSplit: rank 1 nranks 8 total 0.76 (kernels 0.00, alloc 0.00, bootstrap 0.00, allgathers 0.00, topo 0.03, graphs 0.00, connections 0.06, rest 0.67)
|
||
ywang29-p4d-debug-2-worker-0:83698:84146 [2] NCCL INFO Init timings - ncclCommSplit: rank 2 nranks 8 total 0.44 (kernels 0.00, alloc 0.00, bootstrap 0.00, allgathers 0.00, topo 0.03, graphs 0.00, connections 0.06, rest 0.35)
|
||
ywang29-p4d-debug-2-worker-0:83700:84137 [4] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83702:84140 [6] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83701:84134 [5] NCCL INFO NET/OFI NCCL_OFI_TUNER is not available for platform : p4de.24xlarge, Fall back to NCCL's tuner
|
||
ywang29-p4d-debug-2-worker-0:83700:84137 [4] NCCL INFO ncclCommSplit comm 0x562ca54bbc00 rank 4 nranks 8 cudaDev 4 nvmlDev 4 busId 901c0 parent 0x562c99db2980 splitCount 1 color 1266629538 key 4 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83702:84140 [6] NCCL INFO ncclCommSplit comm 0x55a9dbb4c8b0 rank 6 nranks 8 cudaDev 6 nvmlDev 6 busId a01c0 parent 0x55a9d0483a80 splitCount 1 color 1266629538 key 6 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO ncclCommSplit comm 0x55ad937092f0 rank 0 nranks 8 cudaDev 0 nvmlDev 0 busId 101c0 parent 0x55ad880448e0 splitCount 1 color 1266629538 key 0 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83701:84134 [5] NCCL INFO ncclCommSplit comm 0x55dd0ac0a840 rank 5 nranks 8 cudaDev 5 nvmlDev 5 busId 901d0 parent 0x55dcff54b080 splitCount 1 color 1266629538 key 5 - Init COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83700:84137 [4] NCCL INFO Init timings - ncclCommSplit: rank 4 nranks 8 total 0.68 (kernels 0.00, alloc 0.00, bootstrap 0.00, allgathers 0.00, topo 0.03, graphs 0.00, connections 0.06, rest 0.58)
|
||
ywang29-p4d-debug-2-worker-0:83702:84140 [6] NCCL INFO Init timings - ncclCommSplit: rank 6 nranks 8 total 0.66 (kernels 0.00, alloc 0.00, bootstrap 0.00, allgathers 0.00, topo 0.03, graphs 0.00, connections 0.07, rest 0.57)
|
||
ywang29-p4d-debug-2-worker-0:83701:84134 [5] NCCL INFO Init timings - ncclCommSplit: rank 5 nranks 8 total 0.73 (kernels 0.00, alloc 0.00, bootstrap 0.00, allgathers 0.00, topo 0.03, graphs 0.00, connections 0.06, rest 0.63)
|
||
ywang29-p4d-debug-2-worker-0:83696:84149 [0] NCCL INFO Init timings - ncclCommSplit: rank 0 nranks 8 total 0.13 (kernels 0.00, alloc 0.00, bootstrap 0.00, allgathers 0.00, topo 0.03, graphs 0.00, connections 0.06, rest 0.03)
|
||
Stage 3 initialize beginning
|
||
MA 3.44 GB Max_MA 6.34 GB CA 3.48 GB Max_CA 7 GB
|
||
CPU Virtual Memory: used = 54.22 GB, percent = 4.8%
|
||
DeepSpeedZeRoOffload initialize [begin]
|
||
MA 3.44 GB Max_MA 3.44 GB CA 3.48 GB Max_CA 3 GB
|
||
CPU Virtual Memory: used = 54.22 GB, percent = 4.8%
|
||
Parameter Offload - Persistent parameters statistics: param_count = 161, numel = 424960
|
||
DeepSpeedZeRoOffload initialize [end]
|
||
MA 3.44 GB Max_MA 3.44 GB CA 3.48 GB Max_CA 3 GB
|
||
CPU Virtual Memory: used = 54.23 GB, percent = 4.8%
|
||
Before creating fp16 partitions
|
||
MA 3.44 GB Max_MA 3.44 GB CA 3.48 GB Max_CA 3 GB
|
||
CPU Virtual Memory: used = 54.23 GB, percent = 4.8%
|
||
After creating fp16 partitions: 3
|
||
MA 3.44 GB Max_MA 3.44 GB CA 3.46 GB Max_CA 3 GB
|
||
CPU Virtual Memory: used = 54.36 GB, percent = 4.8%
|
||
Before creating fp32 partitions
|
||
MA 3.44 GB Max_MA 3.44 GB CA 3.46 GB Max_CA 3 GB
|
||
CPU Virtual Memory: used = 54.36 GB, percent = 4.8%
|
||
After creating fp32 partitions
|
||
MA 10.32 GB Max_MA 14.06 GB CA 14.06 GB Max_CA 14 GB
|
||
CPU Virtual Memory: used = 54.37 GB, percent = 4.8%
|
||
Before initializing optimizer states
|
||
MA 10.32 GB Max_MA 10.32 GB CA 14.06 GB Max_CA 14 GB
|
||
CPU Virtual Memory: used = 54.38 GB, percent = 4.8%
|
||
After initializing optimizer states
|
||
MA 10.32 GB Max_MA 14.06 GB CA 14.69 GB Max_CA 15 GB
|
||
CPU Virtual Memory: used = 54.38 GB, percent = 4.8%
|
||
After initializing ZeRO optimizer
|
||
MA 13.8 GB Max_MA 16.7 GB CA 17.29 GB Max_CA 17 GB
|
||
CPU Virtual Memory: used = 55.12 GB, percent = 4.9%
|
||
[INFO|trainer.py:1587] 2026-04-30 16:20:08,679 >> ***** Running training *****
|
||
[INFO|trainer.py:1588] 2026-04-30 16:20:08,679 >> Num examples = 30,000
|
||
[INFO|trainer.py:1589] 2026-04-30 16:20:08,679 >> Num Epochs = 4
|
||
[INFO|trainer.py:1590] 2026-04-30 16:20:08,679 >> Instantaneous batch size per device = 1
|
||
[INFO|trainer.py:1593] 2026-04-30 16:20:08,680 >> Total train batch size (w. parallel, distributed & accumulation) = 128
|
||
[INFO|trainer.py:1594] 2026-04-30 16:20:08,680 >> Gradient Accumulation steps = 16
|
||
[INFO|trainer.py:1595] 2026-04-30 16:20:08,680 >> Total optimization steps = 940
|
||
[INFO|trainer.py:1596] 2026-04-30 16:20:08,681 >> Number of trainable parameters = 14,768,307,200
|
||
wandb: [wandb.login()] Loaded credentials for https://api.wandb.ai from WANDB_API_KEY.
|
||
wandb: Currently logged in as: kkhya (maskmoe) to https://api.wandb.ai. Use `wandb login --relogin` to force relogin
|
||
wandb: Tracking run with wandb version 0.26.1
|
||
wandb: Run data is saved locally in /nfs/ywang29/lm-factory/wandb/run-20260430_162009-4wgsdvid
|
||
wandb: Run `wandb offline` to turn off syncing.
|
||
wandb: Syncing run qwen3_14b_coding_fft
|
||
wandb: ⭐️ View project at https://wandb.ai/maskmoe/MFT-LM
|
||
wandb: 🚀 View run at https://wandb.ai/maskmoe/MFT-LM/runs/4wgsdvid
|
||
0%| | 0/940 [00:00<?, ?it/s]ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 00/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 00/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 00/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 00/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 01/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 01/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 01/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 01/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 02/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 02/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 02/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 03/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 03/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 03/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 02/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 03/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 04/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 04/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 05/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 00/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 00/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 04/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 04/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 05/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 01/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 01/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 05/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 05/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 06/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 02/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 02/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 06/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 06/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 06/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 07/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 07/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 07/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 03/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 07/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 08/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 08/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 04/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 08/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 03/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 08/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 09/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 09/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 09/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 04/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 10/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 09/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 05/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 10/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 11/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 05/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 10/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 10/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 11/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 06/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 12/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 07/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 12/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 11/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 11/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 12/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 12/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 08/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 13/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 13/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 06/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 13/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 14/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 14/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 09/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 13/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 10/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 07/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 00/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 15/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 15/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 14/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 14/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 08/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 16/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 16/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 01/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 15/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 16/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 17/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 17/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 11/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 02/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 03/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 15/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 17/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 12/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 18/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 18/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 09/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 10/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 04/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 19/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 19/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 13/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 14/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 16/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 20/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 20/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 11/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 15/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 17/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 12/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 18/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 05/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 21/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 21/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 19/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 06/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 22/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 16/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 22/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Channel 23/0 : 2[2] -> 3[3] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 13/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 18/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Channel 23/0 : 6[6] -> 7[7] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 20/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 17/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 07/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 14/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 21/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 08/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 18/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 19/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 19/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 22/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 09/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 15/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 20/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 10/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Channel 23/0 : 5[5] -> 6[6] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 20/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 16/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 11/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 21/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 17/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 12/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 21/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 22/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 13/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 22/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Channel 23/0 : 0[0] -> 1[1] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 14/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 18/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Channel 23/0 : 3[3] -> 4[4] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 15/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 19/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 16/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 20/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 17/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 21/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 22/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 18/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Channel 23/0 : 4[4] -> 5[5] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 19/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 20/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 21/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 22/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Channel 23/0 : 1[1] -> 2[2] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 00/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 01/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 02/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 03/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 04/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 05/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 06/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 07/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 08/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 09/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 10/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 11/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 12/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 13/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 14/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 15/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 16/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 17/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 18/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 19/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 20/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 21/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 22/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Channel 23/0 : 7[7] -> 0[0] via P2P/CUMEM/read
|
||
ywang29-p4d-debug-2-worker-0:83702:84511 [6] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
ywang29-p4d-debug-2-worker-0:83698:84512 [2] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
ywang29-p4d-debug-2-worker-0:83697:84515 [1] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
ywang29-p4d-debug-2-worker-0:83701:84509 [5] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
ywang29-p4d-debug-2-worker-0:83696:84510 [0] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
ywang29-p4d-debug-2-worker-0:83703:84516 [7] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
ywang29-p4d-debug-2-worker-0:83700:84514 [4] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
ywang29-p4d-debug-2-worker-0:83699:84513 [3] NCCL INFO Connected all rings, use ring PXN 0 GDR 1
|
||
0%| | 1/940 [00:22<5:45:22, 22.07s/it]
0%| | 2/940 [00:41<5:19:36, 20.44s/it]
0%| | 3/940 [00:59<5:00:53, 19.27s/it]
0%| | 4/940 [01:18<4:57:28, 19.07s/it]
1%| | 5/940 [01:38<5:02:43, 19.43s/it]
1%| | 6/940 [01:57<5:01:04, 19.34s/it]
1%| | 7/940 [02:17<5:05:34, 19.65s/it]
1%| | 8/940 [02:38<5:13:56, 20.21s/it]
1%| | 9/940 [02:59<5:14:40, 20.28s/it]
1%| | 10/940 [03:18<5:10:38, 20.04s/it]
{'loss': '0.9603', 'grad_norm': '9', 'learning_rate': '9.574e-07', 'epoch': '0.04267'}
|
||
1%| | 10/940 [03:18<5:10:38, 20.04s/it]
1%| | 11/940 [03:38<5:10:24, 20.05s/it]
1%|▏ | 12/940 [03:59<5:13:58, 20.30s/it]
1%|▏ | 13/940 [04:19<5:10:41, 20.11s/it]
1%|▏ | 14/940 [04:38<5:05:55, 19.82s/it]
2%|▏ | 15/940 [04:58<5:06:45, 19.90s/it]
2%|▏ | 16/940 [05:18<5:07:56, 20.00s/it]
2%|▏ | 17/940 [05:38<5:05:07, 19.83s/it]
2%|▏ | 18/940 [05:58<5:07:58, 20.04s/it]
2%|▏ | 19/940 [06:18<5:06:07, 19.94s/it]
2%|▏ | 20/940 [06:38<5:05:35, 19.93s/it]
{'loss': '0.8527', 'grad_norm': '2.637', 'learning_rate': '2.021e-06', 'epoch': '0.08533'}
|
||
2%|▏ | 20/940 [06:38<5:05:35, 19.93s/it]
2%|▏ | 21/940 [07:01<5:18:19, 20.78s/it]
2%|▏ | 22/940 [07:21<5:14:26, 20.55s/it]
2%|▏ | 23/940 [07:42<5:14:44, 20.59s/it]
3%|▎ | 24/940 [08:00<5:04:18, 19.93s/it]
3%|▎ | 25/940 [08:20<5:02:48, 19.86s/it]
3%|▎ | 26/940 [08:41<5:11:34, 20.45s/it]
3%|▎ | 27/940 [09:01<5:07:04, 20.18s/it]
3%|▎ | 28/940 [09:19<4:57:45, 19.59s/it]
3%|▎ | 29/940 [09:37<4:48:37, 19.01s/it]
3%|▎ | 30/940 [09:54<4:41:51, 18.58s/it]
{'loss': '0.7081', 'grad_norm': '0.9866', 'learning_rate': '3.085e-06', 'epoch': '0.128'}
|
||
3%|▎ | 30/940 [09:54<4:41:51, 18.58s/it]
3%|▎ | 31/940 [10:13<4:39:35, 18.45s/it]
3%|▎ | 32/940 [10:31<4:39:15, 18.45s/it]
4%|▎ | 33/940 [10:50<4:39:34, 18.49s/it]
4%|▎ | 34/940 [11:08<4:39:52, 18.53s/it]
4%|▎ | 35/940 [11:29<4:48:02, 19.10s/it]
4%|▍ | 36/940 [11:47<4:44:22, 18.87s/it]
4%|▍ | 37/940 [12:09<4:56:22, 19.69s/it]
4%|▍ | 38/940 [12:28<4:55:17, 19.64s/it]
4%|▍ | 39/940 [12:46<4:45:58, 19.04s/it]
4%|▍ | 40/940 [13:04<4:41:22, 18.76s/it]
{'loss': '0.6635', 'grad_norm': '0.7401', 'learning_rate': '4.149e-06', 'epoch': '0.1707'}
|
||
4%|▍ | 40/940 [13:04<4:41:22, 18.76s/it]
4%|▍ | 41/940 [13:24<4:48:56, 19.28s/it]
4%|▍ | 42/940 [13:45<4:52:55, 19.57s/it]
5%|▍ | 43/940 [14:07<5:03:00, 20.27s/it]
5%|▍ | 44/940 [14:28<5:08:12, 20.64s/it]
5%|▍ | 45/940 [14:48<5:06:45, 20.57s/it]
5%|▍ | 46/940 [15:10<5:09:46, 20.79s/it]
5%|▌ | 47/940 [15:29<5:00:28, 20.19s/it]
5%|▌ | 48/940 [15:50<5:04:09, 20.46s/it]
5%|▌ | 49/940 [16:08<4:53:14, 19.75s/it]
5%|▌ | 50/940 [16:27<4:49:26, 19.51s/it]
{'loss': '0.6684', 'grad_norm': '0.6697', 'learning_rate': '5.213e-06', 'epoch': '0.2133'}
|
||
5%|▌ | 50/940 [16:27<4:49:26, 19.51s/it]
5%|▌ | 51/940 [16:45<4:42:18, 19.05s/it]
6%|▌ | 52/940 [17:05<4:46:06, 19.33s/it]
6%|▌ | 53/940 [17:23<4:41:07, 19.02s/it]
6%|▌ | 54/940 [17:41<4:35:45, 18.67s/it]
6%|▌ | 55/940 [18:00<4:36:35, 18.75s/it]
6%|▌ | 56/940 [18:21<4:45:28, 19.38s/it]
6%|▌ | 57/940 [18:38<4:36:57, 18.82s/it]
6%|▌ | 58/940 [18:56<4:32:56, 18.57s/it]
6%|▋ | 59/940 [19:16<4:36:35, 18.84s/it]
6%|▋ | 60/940 [19:38<4:53:11, 19.99s/it]
{'loss': '0.6275', 'grad_norm': '0.636', 'learning_rate': '6.277e-06', 'epoch': '0.256'}
|
||
6%|▋ | 60/940 [19:38<4:53:11, 19.99s/it]
6%|▋ | 61/940 [20:00<4:59:09, 20.42s/it]
7%|▋ | 62/940 [20:22<5:07:18, 21.00s/it]
7%|▋ | 63/940 [20:45<5:16:16, 21.64s/it]
7%|▋ | 64/940 [21:10<5:29:03, 22.54s/it]
7%|▋ | 65/940 [21:28<5:09:57, 21.25s/it]
7%|▋ | 66/940 [21:47<5:01:14, 20.68s/it]
7%|▋ | 67/940 [22:07<4:55:39, 20.32s/it]
7%|▋ | 68/940 [22:25<4:47:52, 19.81s/it]
7%|▋ | 69/940 [22:48<4:58:48, 20.58s/it]
7%|▋ | 70/940 [23:08<4:57:01, 20.48s/it]
{'loss': '0.6529', 'grad_norm': '0.5553', 'learning_rate': '7.34e-06', 'epoch': '0.2987'}
|
||
7%|▋ | 70/940 [23:08<4:57:01, 20.48s/it]
8%|▊ | 71/940 [23:27<4:50:17, 20.04s/it]
8%|▊ | 72/940 [23:47<4:47:36, 19.88s/it]
8%|▊ | 73/940 [24:04<4:38:13, 19.25s/it]
8%|▊ | 74/940 [24:23<4:34:56, 19.05s/it]
8%|▊ | 75/940 [24:43<4:40:46, 19.48s/it]
8%|▊ | 76/940 [25:05<4:50:02, 20.14s/it]
8%|▊ | 77/940 [25:24<4:43:34, 19.72s/it]
8%|▊ | 78/940 [25:42<4:36:26, 19.24s/it]
8%|▊ | 79/940 [26:01<4:36:33, 19.27s/it]
9%|▊ | 80/940 [26:20<4:32:53, 19.04s/it]
{'loss': '0.644', 'grad_norm': '0.6058', 'learning_rate': '8.404e-06', 'epoch': '0.3413'}
|
||
9%|▊ | 80/940 [26:20<4:32:53, 19.04s/it]
9%|▊ | 81/940 [26:41<4:41:47, 19.68s/it]
9%|▊ | 82/940 [27:04<4:53:35, 20.53s/it]
9%|▉ | 83/940 [27:21<4:40:55, 19.67s/it]
9%|▉ | 84/940 [27:44<4:55:28, 20.71s/it]
9%|▉ | 85/940 [28:04<4:49:46, 20.34s/it]
9%|▉ | 86/940 [28:23<4:44:11, 19.97s/it]
9%|▉ | 87/940 [28:42<4:41:20, 19.79s/it]
9%|▉ | 88/940 [29:02<4:39:36, 19.69s/it]
9%|▉ | 89/940 [29:24<4:49:15, 20.39s/it]
10%|▉ | 90/940 [29:42<4:39:45, 19.75s/it]
{'loss': '0.6391', 'grad_norm': '0.6475', 'learning_rate': '9.468e-06', 'epoch': '0.384'}
|
||
10%|▉ | 90/940 [29:42<4:39:45, 19.75s/it]
10%|▉ | 91/940 [30:03<4:44:52, 20.13s/it]
10%|▉ | 92/940 [30:23<4:43:16, 20.04s/it]
10%|▉ | 93/940 [30:44<4:46:27, 20.29s/it]
10%|█ | 94/940 [31:03<4:41:09, 19.94s/it]
10%|█ | 95/940 [31:22<4:37:33, 19.71s/it]
10%|█ | 96/940 [31:43<4:41:18, 20.00s/it]
10%|█ | 97/940 [32:04<4:45:27, 20.32s/it]
10%|█ | 98/940 [32:26<4:53:02, 20.88s/it]
11%|█ | 99/940 [32:45<4:46:22, 20.43s/it]
11%|█ | 100/940 [33:05<4:43:29, 20.25s/it]
{'loss': '0.6257', 'grad_norm': '0.5742', 'learning_rate': '9.999e-06', 'epoch': '0.4267'}
|
||
11%|█ | 100/940 [33:05<4:43:29, 20.25s/it]
11%|█ | 101/940 [33:27<4:51:23, 20.84s/it]
11%|█ | 102/940 [33:47<4:46:01, 20.48s/it]
11%|█ | 103/940 [34:06<4:41:17, 20.16s/it]
11%|█ | 104/940 [34:25<4:34:56, 19.73s/it]
11%|█ | 105/940 [34:47<4:42:26, 20.30s/it]
11%|█▏ | 106/940 [35:07<4:41:09, 20.23s/it]
11%|█▏ | 107/940 [35:28<4:45:05, 20.54s/it]
11%|█▏ | 108/940 [35:48<4:43:35, 20.45s/it]
12%|█▏ | 109/940 [36:07<4:34:35, 19.83s/it]
12%|█▏ | 110/940 [36:31<4:53:04, 21.19s/it]
{'loss': '0.5927', 'grad_norm': '0.5906', 'learning_rate': '9.992e-06', 'epoch': '0.4693'}
|
||
12%|█▏ | 110/940 [36:31<4:53:04, 21.19s/it]
12%|█▏ | 111/940 [36:51<4:45:24, 20.66s/it]
12%|█▏ | 112/940 [37:09<4:34:40, 19.90s/it]
12%|█▏ | 113/940 [37:27<4:28:41, 19.49s/it]
12%|█▏ | 114/940 [37:46<4:25:27, 19.28s/it]
12%|█▏ | 115/940 [38:06<4:29:31, 19.60s/it]
12%|█▏ | 116/940 [38:26<4:29:26, 19.62s/it]
12%|█▏ | 117/940 [38:48<4:39:42, 20.39s/it]
13%|█▎ | 118/940 [39:07<4:34:52, 20.06s/it]
13%|█▎ | 119/940 [39:28<4:35:26, 20.13s/it]
13%|█▎ | 120/940 [39:50<4:42:23, 20.66s/it]
{'loss': '0.6213', 'grad_norm': '0.5436', 'learning_rate': '9.978e-06', 'epoch': '0.512'}
|
||
13%|█▎ | 120/940 [39:50<4:42:23, 20.66s/it]
13%|█▎ | 121/940 [40:09<4:36:11, 20.23s/it]
13%|█▎ | 122/940 [40:30<4:39:07, 20.47s/it]
13%|█▎ | 123/940 [40:52<4:44:25, 20.89s/it]
13%|█▎ | 124/940 [41:16<4:55:42, 21.74s/it]
13%|█▎ | 125/940 [41:38<4:56:36, 21.84s/it]
13%|█▎ | 126/940 [41:56<4:42:32, 20.83s/it]
14%|█▎ | 127/940 [42:15<4:36:06, 20.38s/it]
14%|█▎ | 128/940 [42:37<4:38:51, 20.61s/it]
14%|█▎ | 129/940 [42:59<4:46:55, 21.23s/it]
14%|█▍ | 130/940 [43:18<4:38:32, 20.63s/it]
{'loss': '0.6226', 'grad_norm': '0.573', 'learning_rate': '9.958e-06', 'epoch': '0.5547'}
|
||
14%|█▍ | 130/940 [43:18<4:38:32, 20.63s/it]
14%|█▍ | 131/940 [43:38<4:32:59, 20.25s/it]
14%|█▍ | 132/940 [44:00<4:42:34, 20.98s/it]
14%|█▍ | 133/940 [44:19<4:32:30, 20.26s/it]
14%|█▍ | 134/940 [44:41<4:38:04, 20.70s/it]
14%|█▍ | 135/940 [45:03<4:44:38, 21.21s/it]
14%|█▍ | 136/940 [45:22<4:36:31, 20.64s/it]
15%|█▍ | 137/940 [45:45<4:41:58, 21.07s/it]
15%|█▍ | 138/940 [46:05<4:37:59, 20.80s/it]
15%|█▍ | 139/940 [46:24<4:31:38, 20.35s/it]
15%|█▍ | 140/940 [46:43<4:24:11, 19.81s/it]
{'loss': '0.619', 'grad_norm': '0.59', 'learning_rate': '9.93e-06', 'epoch': '0.5973'}
|
||
15%|█▍ | 140/940 [46:43<4:24:11, 19.81s/it]
15%|█▌ | 141/940 [47:01<4:17:59, 19.37s/it]
15%|█▌ | 142/940 [47:21<4:18:39, 19.45s/it]
15%|█▌ | 143/940 [47:38<4:11:31, 18.94s/it]
15%|█▌ | 144/940 [47:56<4:07:59, 18.69s/it]
15%|█▌ | 145/940 [48:15<4:07:21, 18.67s/it]
16%|█▌ | 146/940 [48:33<4:03:27, 18.40s/it]
16%|█▌ | 147/940 [48:53<4:10:49, 18.98s/it]
16%|█▌ | 148/940 [49:12<4:09:55, 18.93s/it]
16%|█▌ | 149/940 [49:34<4:20:18, 19.75s/it]
16%|█▌ | 150/940 [49:54<4:21:07, 19.83s/it]
{'loss': '0.6111', 'grad_norm': '0.5526', 'learning_rate': '9.896e-06', 'epoch': '0.64'}
|
||
16%|█▌ | 150/940 [49:54<4:21:07, 19.83s/it]
16%|█▌ | 151/940 [50:12<4:13:50, 19.30s/it]
16%|█▌ | 152/940 [50:31<4:11:41, 19.16s/it]
16%|█▋ | 153/940 [50:50<4:11:45, 19.19s/it]
16%|█▋ | 154/940 [51:09<4:11:23, 19.19s/it]
16%|█▋ | 155/940 [51:33<4:29:48, 20.62s/it]
17%|█▋ | 156/940 [51:53<4:27:10, 20.45s/it]
17%|█▋ | 157/940 [52:14<4:28:18, 20.56s/it]
17%|█▋ | 158/940 [52:36<4:34:01, 21.03s/it]
17%|█▋ | 159/940 [52:58<4:39:09, 21.45s/it]
17%|█▋ | 160/940 [53:18<4:32:48, 20.99s/it]
{'loss': '0.6422', 'grad_norm': '0.5711', 'learning_rate': '9.855e-06', 'epoch': '0.6827'}
|
||
17%|█▋ | 160/940 [53:18<4:32:48, 20.99s/it]
17%|█▋ | 161/940 [53:37<4:22:57, 20.25s/it]
17%|█▋ | 162/940 [53:56<4:18:47, 19.96s/it]
17%|█▋ | 163/940 [54:18<4:25:47, 20.52s/it]
17%|█▋ | 164/940 [54:40<4:32:14, 21.05s/it]
18%|█▊ | 165/940 [55:04<4:41:41, 21.81s/it]
18%|█▊ | 166/940 [55:25<4:38:26, 21.59s/it]
18%|█▊ | 167/940 [55:47<4:38:46, 21.64s/it]
18%|█▊ | 168/940 [56:10<4:43:59, 22.07s/it]
18%|█▊ | 169/940 [56:29<4:33:40, 21.30s/it]
18%|█▊ | 170/940 [56:47<4:21:41, 20.39s/it]
{'loss': '0.6056', 'grad_norm': '0.5616', 'learning_rate': '9.807e-06', 'epoch': '0.7253'}
|
||
18%|█▊ | 170/940 [56:47<4:21:41, 20.39s/it]
18%|█▊ | 171/940 [57:07<4:18:08, 20.14s/it]
18%|█▊ | 172/940 [57:28<4:20:40, 20.37s/it]
18%|█▊ | 173/940 [57:46<4:12:36, 19.76s/it]
19%|█▊ | 174/940 [58:05<4:06:30, 19.31s/it]
19%|█▊ | 175/940 [58:23<4:02:52, 19.05s/it]
19%|█▊ | 176/940 [58:41<3:57:53, 18.68s/it]
19%|█▉ | 177/940 [59:02<4:07:18, 19.45s/it]
19%|█▉ | 178/940 [59:26<4:24:49, 20.85s/it]
19%|█▉ | 179/940 [59:48<4:28:03, 21.13s/it]
19%|█▉ | 180/940 [1:00:08<4:25:27, 20.96s/it]
{'loss': '0.6315', 'grad_norm': '0.5977', 'learning_rate': '9.753e-06', 'epoch': '0.768'}
|
||
19%|█▉ | 180/940 [1:00:08<4:25:27, 20.96s/it]
19%|█▉ | 181/940 [1:00:32<4:34:54, 21.73s/it]
19%|█▉ | 182/940 [1:00:53<4:32:15, 21.55s/it]
19%|█▉ | 183/940 [1:01:13<4:24:30, 20.96s/it]
20%|█▉ | 184/940 [1:01:31<4:13:10, 20.09s/it]
20%|█▉ | 185/940 [1:01:50<4:07:59, 19.71s/it]
20%|█▉ | 186/940 [1:02:09<4:05:17, 19.52s/it]
20%|█▉ | 187/940 [1:02:28<4:03:55, 19.44s/it]
20%|██ | 188/940 [1:02:49<4:08:35, 19.83s/it]
20%|██ | 189/940 [1:03:10<4:11:55, 20.13s/it]
20%|██ | 190/940 [1:03:31<4:18:24, 20.67s/it]
{'loss': '0.6109', 'grad_norm': '0.5179', 'learning_rate': '9.692e-06', 'epoch': '0.8107'}
|
||
20%|██ | 190/940 [1:03:31<4:18:24, 20.67s/it]
20%|██ | 191/940 [1:03:52<4:17:49, 20.65s/it]
20%|██ | 192/940 [1:04:12<4:16:34, 20.58s/it]
21%|██ | 193/940 [1:04:33<4:15:03, 20.49s/it]
21%|██ | 194/940 [1:04:53<4:13:36, 20.40s/it]
21%|██ | 195/940 [1:05:13<4:12:38, 20.35s/it]
21%|██ | 196/940 [1:05:34<4:14:27, 20.52s/it]
21%|██ | 197/940 [1:05:55<4:14:17, 20.53s/it]
21%|██ | 198/940 [1:06:15<4:13:02, 20.46s/it]
21%|██ | 199/940 [1:06:37<4:19:32, 21.02s/it]
21%|██▏ | 200/940 [1:06:56<4:10:03, 20.28s/it]
{'loss': '0.6224', 'grad_norm': '0.5373', 'learning_rate': '9.625e-06', 'epoch': '0.8533'}
|
||
21%|██▏ | 200/940 [1:06:56<4:10:03, 20.28s/it]
21%|██▏ | 201/940 [1:07:17<4:12:04, 20.47s/it]
21%|██▏ | 202/940 [1:07:38<4:16:01, 20.81s/it]
22%|██▏ | 203/940 [1:08:01<4:22:53, 21.40s/it]
22%|██▏ | 204/940 [1:08:22<4:20:04, 21.20s/it]
22%|██▏ | 205/940 [1:08:44<4:24:59, 21.63s/it]
22%|██▏ | 206/940 [1:09:05<4:21:27, 21.37s/it]
22%|██▏ | 207/940 [1:09:23<4:07:32, 20.26s/it]
22%|██▏ | 208/940 [1:09:42<4:02:07, 19.85s/it]
22%|██▏ | 209/940 [1:10:02<4:04:24, 20.06s/it]
22%|██▏ | 210/940 [1:10:24<4:09:51, 20.54s/it]
{'loss': '0.6081', 'grad_norm': '0.5947', 'learning_rate': '9.551e-06', 'epoch': '0.896'}
|
||
22%|██▏ | 210/940 [1:10:24<4:09:51, 20.54s/it]
22%|██▏ | 211/940 [1:10:44<4:08:01, 20.41s/it]
23%|██▎ | 212/940 [1:11:02<3:59:03, 19.70s/it]
23%|██▎ | 213/940 [1:11:21<3:55:21, 19.42s/it]
23%|██▎ | 214/940 [1:11:41<3:57:18, 19.61s/it]
23%|██▎ | 215/940 [1:12:00<3:56:15, 19.55s/it]
23%|██▎ | 216/940 [1:12:24<4:08:52, 20.62s/it]
23%|██▎ | 217/940 [1:12:42<4:01:29, 20.04s/it]
23%|██▎ | 218/940 [1:13:01<3:57:59, 19.78s/it]
23%|██▎ | 219/940 [1:13:20<3:54:39, 19.53s/it]
23%|██▎ | 220/940 [1:13:38<3:48:19, 19.03s/it]
{'loss': '0.6155', 'grad_norm': '0.5535', 'learning_rate': '9.471e-06', 'epoch': '0.9387'}
|
||
23%|██▎ | 220/940 [1:13:38<3:48:19, 19.03s/it]
24%|██▎ | 221/940 [1:13:59<3:52:37, 19.41s/it]
24%|██▎ | 222/940 [1:14:17<3:50:10, 19.24s/it]
24%|██▎ | 223/940 [1:14:38<3:53:24, 19.53s/it]
24%|██▍ | 224/940 [1:14:56<3:49:07, 19.20s/it]
24%|██▍ | 225/940 [1:15:13<3:42:47, 18.70s/it]
24%|██▍ | 226/940 [1:15:32<3:41:11, 18.59s/it]
24%|██▍ | 227/940 [1:15:50<3:38:28, 18.39s/it]
24%|██▍ | 228/940 [1:16:08<3:36:35, 18.25s/it]
24%|██▍ | 229/940 [1:16:26<3:35:22, 18.17s/it]
24%|██▍ | 230/940 [1:16:47<3:47:24, 19.22s/it]
{'loss': '0.6901', 'grad_norm': '0.6431', 'learning_rate': '9.385e-06', 'epoch': '0.9813'}
|
||
24%|██▍ | 230/940 [1:16:47<3:47:24, 19.22s/it]
25%|██▍ | 231/940 [1:17:05<3:42:11, 18.80s/it]
25%|██▍ | 232/940 [1:17:25<3:45:38, 19.12s/it]
25%|██▍ | 233/940 [1:17:46<3:52:21, 19.72s/it]
25%|██▍ | 234/940 [1:18:06<3:51:36, 19.68s/it]
25%|██▌ | 235/940 [1:18:16<3:16:22, 16.71s/it]
25%|██▌ | 236/940 [1:18:38<3:34:48, 18.31s/it]
25%|██▌ | 237/940 [1:19:00<3:47:27, 19.41s/it]
25%|██▌ | 238/940 [1:19:20<3:49:39, 19.63s/it]
25%|██▌ | 239/940 [1:19:44<4:04:20, 20.91s/it]
26%|██▌ | 240/940 [1:20:02<3:56:33, 20.28s/it]
{'loss': '0.5848', 'grad_norm': '0.5828', 'learning_rate': '9.293e-06', 'epoch': '1.021'}
|
||
26%|██▌ | 240/940 [1:20:02<3:56:33, 20.28s/it]
26%|██▌ | 241/940 [1:20:22<3:54:04, 20.09s/it]
26%|██▌ | 242/940 [1:20:40<3:46:59, 19.51s/it]
26%|██▌ | 243/940 [1:21:01<3:49:33, 19.76s/it]
26%|██▌ | 244/940 [1:21:19<3:43:13, 19.24s/it]
26%|██▌ | 245/940 [1:21:42<3:57:26, 20.50s/it]
26%|██▌ | 246/940 [1:22:01<3:53:04, 20.15s/it]
26%|██▋ | 247/940 [1:22:27<4:10:44, 21.71s/it]
26%|██▋ | 248/940 [1:22:47<4:05:37, 21.30s/it]
26%|██▋ | 249/940 [1:23:08<4:03:28, 21.14s/it]
27%|██▋ | 250/940 [1:23:29<4:02:00, 21.04s/it]
{'loss': '0.4799', 'grad_norm': '0.5427', 'learning_rate': '9.194e-06', 'epoch': '1.064'}
|
||
27%|██▋ | 250/940 [1:23:29<4:02:00, 21.04s/it]
27%|██▋ | 251/940 [1:23:50<4:01:11, 21.00s/it]
27%|██▋ | 252/940 [1:24:11<4:01:02, 21.02s/it]
27%|██▋ | 253/940 [1:24:31<3:59:13, 20.89s/it]
27%|██▋ | 254/940 [1:24:52<3:59:00, 20.90s/it]
27%|██▋ | 255/940 [1:25:14<4:02:04, 21.20s/it]
27%|██▋ | 256/940 [1:25:40<4:16:28, 22.50s/it]
27%|██▋ | 257/940 [1:26:00<4:09:57, 21.96s/it]
27%|██▋ | 258/940 [1:26:21<4:04:35, 21.52s/it]
28%|██▊ | 259/940 [1:26:43<4:07:19, 21.79s/it]
28%|██▊ | 260/940 [1:27:04<4:04:30, 21.57s/it]
{'loss': '0.503', 'grad_norm': '0.5644', 'learning_rate': '9.09e-06', 'epoch': '1.107'}
|
||
28%|██▊ | 260/940 [1:27:04<4:04:30, 21.57s/it]
28%|██▊ | 261/940 [1:27:22<3:52:23, 20.53s/it]
28%|██▊ | 262/940 [1:27:42<3:50:11, 20.37s/it]
28%|██▊ | 263/940 [1:28:01<3:45:09, 19.96s/it]
28%|██▊ | 264/940 [1:28:24<3:53:03, 20.68s/it]
28%|██▊ | 265/940 [1:28:44<3:51:24, 20.57s/it]
28%|██▊ | 266/940 [1:29:02<3:43:52, 19.93s/it]
28%|██▊ | 267/940 [1:29:23<3:44:41, 20.03s/it]
29%|██▊ | 268/940 [1:29:42<3:41:37, 19.79s/it]
29%|██▊ | 269/940 [1:30:05<3:52:03, 20.75s/it]
29%|██▊ | 270/940 [1:30:27<3:57:08, 21.24s/it]
{'loss': '0.5118', 'grad_norm': '0.6399', 'learning_rate': '8.981e-06', 'epoch': '1.149'}
|
||
29%|██▊ | 270/940 [1:30:27<3:57:08, 21.24s/it]
29%|██▉ | 271/940 [1:30:47<3:50:26, 20.67s/it]
29%|██▉ | 272/940 [1:31:06<3:45:04, 20.22s/it]
29%|██▉ | 273/940 [1:31:24<3:38:15, 19.63s/it]
29%|██▉ | 274/940 [1:31:45<3:43:44, 20.16s/it]
29%|██▉ | 275/940 [1:32:03<3:35:29, 19.44s/it]
29%|██▉ | 276/940 [1:32:22<3:32:55, 19.24s/it]
29%|██▉ | 277/940 [1:32:42<3:33:43, 19.34s/it]
30%|██▉ | 278/940 [1:33:00<3:28:45, 18.92s/it]
30%|██▉ | 279/940 [1:33:18<3:28:22, 18.91s/it]
30%|██▉ | 280/940 [1:33:37<3:26:19, 18.76s/it]
{'loss': '0.5203', 'grad_norm': '0.5862', 'learning_rate': '8.866e-06', 'epoch': '1.192'}
|
||
30%|██▉ | 280/940 [1:33:37<3:26:19, 18.76s/it]
30%|██▉ | 281/940 [1:33:57<3:30:48, 19.19s/it]
30%|███ | 282/940 [1:34:19<3:38:39, 19.94s/it]
30%|███ | 283/940 [1:34:38<3:34:54, 19.63s/it]
30%|███ | 284/940 [1:34:59<3:38:50, 20.02s/it]
30%|███ | 285/940 [1:35:19<3:39:15, 20.08s/it]
30%|███ | 286/940 [1:35:38<3:37:13, 19.93s/it]
31%|███ | 287/940 [1:35:56<3:28:47, 19.18s/it]
31%|███ | 288/940 [1:36:18<3:38:24, 20.10s/it]
31%|███ | 289/940 [1:36:37<3:36:05, 19.92s/it]
31%|███ | 290/940 [1:36:57<3:34:56, 19.84s/it]
{'loss': '0.4948', 'grad_norm': '0.5441', 'learning_rate': '8.745e-06', 'epoch': '1.235'}
|
||
31%|███ | 290/940 [1:36:57<3:34:56, 19.84s/it]
31%|███ | 291/940 [1:37:16<3:32:44, 19.67s/it]
31%|███ | 292/940 [1:37:37<3:34:56, 19.90s/it]
31%|███ | 293/940 [1:37:58<3:39:23, 20.35s/it]
31%|███▏ | 294/940 [1:38:16<3:32:02, 19.69s/it]
31%|███▏ | 295/940 [1:38:39<3:40:24, 20.50s/it]
31%|███▏ | 296/940 [1:38:57<3:33:26, 19.89s/it]
32%|███▏ | 297/940 [1:39:16<3:28:05, 19.42s/it]
32%|███▏ | 298/940 [1:39:34<3:24:08, 19.08s/it]
32%|███▏ | 299/940 [1:39:53<3:22:25, 18.95s/it]
32%|███▏ | 300/940 [1:40:10<3:18:46, 18.63s/it]
{'loss': '0.5043', 'grad_norm': '0.6002', 'learning_rate': '8.62e-06', 'epoch': '1.277'}
|
||
32%|███▏ | 300/940 [1:40:10<3:18:46, 18.63s/it]
32%|███▏ | 301/940 [1:40:31<3:24:29, 19.20s/it]
32%|███▏ | 302/940 [1:40:53<3:33:03, 20.04s/it]
32%|███▏ | 303/940 [1:41:13<3:33:27, 20.11s/it]
32%|███▏ | 304/940 [1:41:36<3:41:06, 20.86s/it]
32%|███▏ | 305/940 [1:41:57<3:41:46, 20.96s/it]
33%|███▎ | 306/940 [1:42:21<3:51:31, 21.91s/it]
33%|███▎ | 307/940 [1:42:40<3:43:04, 21.15s/it]
33%|███▎ | 308/940 [1:43:01<3:40:04, 20.89s/it]
33%|███▎ | 309/940 [1:43:22<3:41:59, 21.11s/it]
33%|███▎ | 310/940 [1:43:44<3:42:55, 21.23s/it]
{'loss': '0.5076', 'grad_norm': '0.5812', 'learning_rate': '8.489e-06', 'epoch': '1.32'}
|
||
33%|███▎ | 310/940 [1:43:44<3:42:55, 21.23s/it]
33%|███▎ | 311/940 [1:44:02<3:32:39, 20.29s/it]
33%|███▎ | 312/940 [1:44:26<3:43:07, 21.32s/it]
33%|███▎ | 313/940 [1:44:43<3:31:06, 20.20s/it]
33%|███▎ | 314/940 [1:45:03<3:30:18, 20.16s/it]
34%|███▎ | 315/940 [1:45:23<3:28:15, 19.99s/it]
34%|███▎ | 316/940 [1:45:42<3:24:33, 19.67s/it]
34%|███▎ | 317/940 [1:46:00<3:18:42, 19.14s/it]
34%|███▍ | 318/940 [1:46:20<3:22:33, 19.54s/it]
34%|███▍ | 319/940 [1:46:39<3:19:08, 19.24s/it]
34%|███▍ | 320/940 [1:46:59<3:20:16, 19.38s/it]
{'loss': '0.5113', 'grad_norm': '0.5388', 'learning_rate': '8.354e-06', 'epoch': '1.363'}
|
||
34%|███▍ | 320/940 [1:46:59<3:20:16, 19.38s/it]
34%|███▍ | 321/940 [1:47:18<3:19:31, 19.34s/it]
34%|███▍ | 322/940 [1:47:36<3:17:13, 19.15s/it]
34%|███▍ | 323/940 [1:47:59<3:26:06, 20.04s/it]
34%|███▍ | 324/940 [1:48:16<3:18:35, 19.34s/it]
35%|███▍ | 325/940 [1:48:39<3:28:27, 20.34s/it]
35%|███▍ | 326/940 [1:48:58<3:22:46, 19.82s/it]
35%|███▍ | 327/940 [1:49:19<3:27:10, 20.28s/it]
35%|███▍ | 328/940 [1:49:38<3:23:39, 19.97s/it]
35%|███▌ | 329/940 [1:49:56<3:16:59, 19.34s/it]
35%|███▌ | 330/940 [1:50:19<3:26:12, 20.28s/it]
{'loss': '0.5259', 'grad_norm': '0.6016', 'learning_rate': '8.214e-06', 'epoch': '1.405'}
|
||
35%|███▌ | 330/940 [1:50:19<3:26:12, 20.28s/it]
35%|███▌ | 331/940 [1:50:40<3:29:44, 20.66s/it]
35%|███▌ | 332/940 [1:51:00<3:25:46, 20.31s/it]
35%|███▌ | 333/940 [1:51:18<3:21:18, 19.90s/it]
36%|███▌ | 334/940 [1:51:37<3:16:12, 19.43s/it]
36%|███▌ | 335/940 [1:51:59<3:24:30, 20.28s/it]
36%|███▌ | 336/940 [1:52:20<3:24:51, 20.35s/it]
36%|███▌ | 337/940 [1:52:40<3:25:59, 20.50s/it]
36%|███▌ | 338/940 [1:53:00<3:22:04, 20.14s/it]
36%|███▌ | 339/940 [1:53:18<3:16:32, 19.62s/it]
36%|███▌ | 340/940 [1:53:37<3:12:44, 19.27s/it]
{'loss': '0.4648', 'grad_norm': '0.5889', 'learning_rate': '8.07e-06', 'epoch': '1.448'}
|
||
36%|███▌ | 340/940 [1:53:37<3:12:44, 19.27s/it]
36%|███▋ | 341/940 [1:53:55<3:10:11, 19.05s/it]
36%|███▋ | 342/940 [1:54:14<3:07:57, 18.86s/it]
36%|███▋ | 343/940 [1:54:33<3:08:19, 18.93s/it]
37%|███▋ | 344/940 [1:54:52<3:10:00, 19.13s/it]
37%|███▋ | 345/940 [1:55:11<3:09:49, 19.14s/it]
37%|███▋ | 346/940 [1:55:30<3:08:00, 18.99s/it]
37%|███▋ | 347/940 [1:55:50<3:09:50, 19.21s/it]
37%|███▋ | 348/940 [1:56:08<3:05:10, 18.77s/it]
37%|███▋ | 349/940 [1:56:28<3:09:45, 19.27s/it]
37%|███▋ | 350/940 [1:56:50<3:17:33, 20.09s/it]
{'loss': '0.5107', 'grad_norm': '0.6118', 'learning_rate': '7.921e-06', 'epoch': '1.491'}
|
||
37%|███▋ | 350/940 [1:56:50<3:17:33, 20.09s/it]
37%|███▋ | 351/940 [1:57:08<3:11:17, 19.49s/it]
37%|███▋ | 352/940 [1:57:30<3:17:07, 20.11s/it]
38%|███▊ | 353/940 [1:57:48<3:10:24, 19.46s/it]
38%|███▊ | 354/940 [1:58:06<3:05:42, 19.01s/it]
38%|███▊ | 355/940 [1:58:25<3:06:23, 19.12s/it]
38%|███▊ | 356/940 [1:58:43<3:04:33, 18.96s/it]
38%|███▊ | 357/940 [1:59:03<3:05:09, 19.06s/it]
38%|███▊ | 358/940 [1:59:24<3:10:33, 19.65s/it]
38%|███▊ | 359/940 [1:59:44<3:12:22, 19.87s/it]
38%|███▊ | 360/940 [2:00:03<3:08:40, 19.52s/it]
{'loss': '0.4999', 'grad_norm': '0.6314', 'learning_rate': '7.768e-06', 'epoch': '1.533'}
|
||
38%|███▊ | 360/940 [2:00:03<3:08:40, 19.52s/it]
38%|███▊ | 361/940 [2:00:25<3:14:41, 20.17s/it]
39%|███▊ | 362/940 [2:00:49<3:27:16, 21.52s/it]
39%|███▊ | 363/940 [2:01:09<3:23:10, 21.13s/it]
39%|███▊ | 364/940 [2:01:30<3:21:41, 21.01s/it]
39%|███▉ | 365/940 [2:01:51<3:21:10, 20.99s/it]
39%|███▉ | 366/940 [2:02:09<3:13:06, 20.19s/it]
39%|███▉ | 367/940 [2:02:28<3:08:42, 19.76s/it]
39%|███▉ | 368/940 [2:02:51<3:16:35, 20.62s/it]
39%|███▉ | 369/940 [2:03:11<3:14:11, 20.41s/it]
39%|███▉ | 370/940 [2:03:31<3:13:49, 20.40s/it]
{'loss': '0.5182', 'grad_norm': '0.6', 'learning_rate': '7.612e-06', 'epoch': '1.576'}
|
||
39%|███▉ | 370/940 [2:03:31<3:13:49, 20.40s/it]
39%|███▉ | 371/940 [2:03:53<3:16:36, 20.73s/it]
40%|███▉ | 372/940 [2:04:11<3:09:35, 20.03s/it]
40%|███▉ | 373/940 [2:04:29<3:04:28, 19.52s/it]
40%|███▉ | 374/940 [2:04:48<3:02:09, 19.31s/it]
40%|███▉ | 375/940 [2:05:08<3:03:45, 19.51s/it]
40%|████ | 376/940 [2:05:29<3:08:30, 20.05s/it]
40%|████ | 377/940 [2:05:49<3:07:35, 19.99s/it]
40%|████ | 378/940 [2:06:08<3:03:42, 19.61s/it]
40%|████ | 379/940 [2:06:27<3:00:18, 19.28s/it]
40%|████ | 380/940 [2:06:48<3:06:32, 19.99s/it]
{'loss': '0.5034', 'grad_norm': '0.5932', 'learning_rate': '7.452e-06', 'epoch': '1.619'}
|
||
40%|████ | 380/940 [2:06:48<3:06:32, 19.99s/it]
41%|████ | 381/940 [2:07:08<3:05:14, 19.88s/it]
41%|████ | 382/940 [2:07:29<3:07:35, 20.17s/it]
41%|████ | 383/940 [2:07:48<3:03:42, 19.79s/it]
41%|████ | 384/940 [2:08:06<2:59:45, 19.40s/it]
41%|████ | 385/940 [2:08:27<3:03:08, 19.80s/it]
41%|████ | 386/940 [2:08:45<2:58:07, 19.29s/it]
41%|████ | 387/940 [2:09:04<2:56:07, 19.11s/it]
41%|████▏ | 388/940 [2:09:22<2:54:26, 18.96s/it]
41%|████▏ | 389/940 [2:09:46<3:06:34, 20.32s/it]
41%|████▏ | 390/940 [2:10:04<3:00:34, 19.70s/it]
{'loss': '0.5124', 'grad_norm': '0.6266', 'learning_rate': '7.288e-06', 'epoch': '1.661'}
|
||
41%|████▏ | 390/940 [2:10:04<3:00:34, 19.70s/it]
42%|████▏ | 391/940 [2:10:29<3:15:10, 21.33s/it]
42%|████▏ | 392/940 [2:10:49<3:12:15, 21.05s/it]
42%|████▏ | 393/940 [2:11:11<3:13:11, 21.19s/it]
42%|████▏ | 394/940 [2:11:32<3:11:48, 21.08s/it]
42%|████▏ | 395/940 [2:11:52<3:09:12, 20.83s/it]
42%|████▏ | 396/940 [2:12:15<3:13:21, 21.33s/it]
42%|████▏ | 397/940 [2:12:34<3:08:36, 20.84s/it]
42%|████▏ | 398/940 [2:12:53<3:03:31, 20.32s/it]
42%|████▏ | 399/940 [2:13:13<3:01:00, 20.08s/it]
43%|████▎ | 400/940 [2:13:31<2:55:30, 19.50s/it]
{'loss': '0.4883', 'grad_norm': '0.5718', 'learning_rate': '7.122e-06', 'epoch': '1.704'}
|
||
43%|████▎ | 400/940 [2:13:31<2:55:30, 19.50s/it]
43%|████▎ | 401/940 [2:13:52<2:57:53, 19.80s/it]
43%|████▎ | 402/940 [2:14:10<2:54:46, 19.49s/it]
43%|████▎ | 403/940 [2:14:30<2:55:07, 19.57s/it]
43%|████▎ | 404/940 [2:14:50<2:55:59, 19.70s/it]
43%|████▎ | 405/940 [2:15:08<2:51:05, 19.19s/it]
43%|████▎ | 406/940 [2:15:28<2:53:18, 19.47s/it]
43%|████▎ | 407/940 [2:15:48<2:54:08, 19.60s/it]
43%|████▎ | 408/940 [2:16:08<2:53:49, 19.60s/it]
44%|████▎ | 409/940 [2:16:29<2:56:55, 19.99s/it]
44%|████▎ | 410/940 [2:16:46<2:50:31, 19.31s/it]
{'loss': '0.4877', 'grad_norm': '0.6003', 'learning_rate': '6.952e-06', 'epoch': '1.747'}
|
||
44%|████▎ | 410/940 [2:16:46<2:50:31, 19.31s/it]
44%|████▎ | 411/940 [2:17:08<2:55:52, 19.95s/it]
44%|████▍ | 412/940 [2:17:30<3:02:53, 20.78s/it]
44%|████▍ | 413/940 [2:17:51<3:03:05, 20.85s/it]
44%|████▍ | 414/940 [2:18:12<3:00:59, 20.64s/it]
44%|████▍ | 415/940 [2:18:34<3:04:34, 21.09s/it]
44%|████▍ | 416/940 [2:18:55<3:05:07, 21.20s/it]
44%|████▍ | 417/940 [2:19:17<3:07:29, 21.51s/it]
44%|████▍ | 418/940 [2:19:38<3:03:44, 21.12s/it]
45%|████▍ | 419/940 [2:19:59<3:04:42, 21.27s/it]
45%|████▍ | 420/940 [2:20:22<3:07:29, 21.63s/it]
{'loss': '0.5198', 'grad_norm': '0.5937', 'learning_rate': '6.78e-06', 'epoch': '1.789'}
|
||
45%|████▍ | 420/940 [2:20:22<3:07:29, 21.63s/it]
45%|████▍ | 421/940 [2:20:43<3:05:49, 21.48s/it]
45%|████▍ | 422/940 [2:21:04<3:04:34, 21.38s/it]
45%|████▌ | 423/940 [2:21:24<3:01:45, 21.09s/it]
45%|████▌ | 424/940 [2:21:47<3:05:45, 21.60s/it]
45%|████▌ | 425/940 [2:22:09<3:05:32, 21.62s/it]
45%|████▌ | 426/940 [2:22:30<3:04:07, 21.49s/it]
45%|████▌ | 427/940 [2:22:51<3:02:10, 21.31s/it]
46%|████▌ | 428/940 [2:23:16<3:11:04, 22.39s/it]
46%|████▌ | 429/940 [2:23:37<3:06:36, 21.91s/it]
46%|████▌ | 430/940 [2:24:00<3:10:17, 22.39s/it]
{'loss': '0.5153', 'grad_norm': '0.574', 'learning_rate': '6.605e-06', 'epoch': '1.832'}
|
||
46%|████▌ | 430/940 [2:24:00<3:10:17, 22.39s/it]
46%|████▌ | 431/940 [2:24:23<3:10:54, 22.50s/it]
46%|████▌ | 432/940 [2:24:46<3:11:57, 22.67s/it]
46%|████▌ | 433/940 [2:25:07<3:07:08, 22.15s/it]
46%|████▌ | 434/940 [2:25:27<3:02:06, 21.59s/it]
46%|████▋ | 435/940 [2:25:51<3:06:40, 22.18s/it]
46%|████▋ | 436/940 [2:26:17<3:16:11, 23.36s/it]
46%|████▋ | 437/940 [2:26:39<3:12:29, 22.96s/it]
47%|████▋ | 438/940 [2:26:59<3:05:17, 22.15s/it]
47%|████▋ | 439/940 [2:27:17<2:54:37, 20.91s/it]
47%|████▋ | 440/940 [2:27:36<2:50:02, 20.41s/it]
{'loss': '0.5219', 'grad_norm': '0.6292', 'learning_rate': '6.428e-06', 'epoch': '1.875'}
|
||
47%|████▋ | 440/940 [2:27:36<2:50:02, 20.41s/it]
47%|████▋ | 441/940 [2:27:57<2:50:49, 20.54s/it]
47%|████▋ | 442/940 [2:28:16<2:45:19, 19.92s/it]
47%|████▋ | 443/940 [2:28:39<2:53:33, 20.95s/it]
47%|████▋ | 444/940 [2:28:58<2:48:16, 20.36s/it]
47%|████▋ | 445/940 [2:29:21<2:53:55, 21.08s/it]
47%|████▋ | 446/940 [2:29:43<2:55:55, 21.37s/it]
48%|████▊ | 447/940 [2:30:03<2:51:13, 20.84s/it]
48%|████▊ | 448/940 [2:30:20<2:43:50, 19.98s/it]
48%|████▊ | 449/940 [2:30:39<2:38:43, 19.40s/it]
48%|████▊ | 450/940 [2:30:56<2:34:14, 18.89s/it]
{'loss': '0.5165', 'grad_norm': '0.6479', 'learning_rate': '6.249e-06', 'epoch': '1.917'}
|
||
48%|████▊ | 450/940 [2:30:56<2:34:14, 18.89s/it]
48%|████▊ | 451/940 [2:31:17<2:38:10, 19.41s/it]
48%|████▊ | 452/940 [2:31:36<2:36:39, 19.26s/it]
48%|████▊ | 453/940 [2:31:56<2:37:33, 19.41s/it]
48%|████▊ | 454/940 [2:32:17<2:42:58, 20.12s/it]
48%|████▊ | 455/940 [2:32:37<2:41:44, 20.01s/it]
49%|████▊ | 456/940 [2:32:55<2:36:22, 19.39s/it]
49%|████▊ | 457/940 [2:33:14<2:34:49, 19.23s/it]
49%|████▊ | 458/940 [2:33:33<2:34:23, 19.22s/it]
49%|████▉ | 459/940 [2:33:55<2:40:15, 19.99s/it]
49%|████▉ | 460/940 [2:34:15<2:40:35, 20.07s/it]
{'loss': '0.492', 'grad_norm': '0.5571', 'learning_rate': '6.069e-06', 'epoch': '1.96'}
|
||
49%|████▉ | 460/940 [2:34:15<2:40:35, 20.07s/it]
49%|████▉ | 461/940 [2:34:34<2:37:21, 19.71s/it]
49%|████▉ | 462/940 [2:34:54<2:38:22, 19.88s/it]
49%|████▉ | 463/940 [2:35:13<2:36:20, 19.67s/it]
49%|████▉ | 464/940 [2:35:34<2:37:17, 19.83s/it]
49%|████▉ | 465/940 [2:35:54<2:37:17, 19.87s/it]
50%|████▉ | 466/940 [2:36:13<2:34:53, 19.61s/it]
50%|████▉ | 467/940 [2:36:35<2:40:34, 20.37s/it]
50%|████▉ | 468/940 [2:36:56<2:41:37, 20.54s/it]
50%|████▉ | 469/940 [2:37:15<2:38:15, 20.16s/it]
50%|█████ | 470/940 [2:37:22<2:07:15, 16.25s/it]
{'loss': '0.5182', 'grad_norm': '0.9597', 'learning_rate': '5.887e-06', 'epoch': '2'}
|
||
50%|█████ | 470/940 [2:37:22<2:07:15, 16.25s/it]
50%|█████ | 471/940 [2:37:43<2:16:51, 17.51s/it]
50%|█████ | 472/940 [2:38:02<2:20:51, 18.06s/it]
50%|█████ | 473/940 [2:38:22<2:25:51, 18.74s/it]
50%|█████ | 474/940 [2:38:41<2:26:10, 18.82s/it]
51%|█████ | 475/940 [2:39:00<2:24:55, 18.70s/it]
51%|█████ | 476/940 [2:39:17<2:22:15, 18.40s/it]
51%|█████ | 477/940 [2:39:38<2:27:39, 19.13s/it]
51%|█████ | 478/940 [2:39:59<2:30:13, 19.51s/it]
51%|█████ | 479/940 [2:40:17<2:27:20, 19.18s/it]
51%|█████ | 480/940 [2:40:37<2:28:57, 19.43s/it]
{'loss': '0.3669', 'grad_norm': '0.7599', 'learning_rate': '5.703e-06', 'epoch': '2.043'}
|
||
51%|█████ | 480/940 [2:40:37<2:28:57, 19.43s/it]
51%|█████ | 481/940 [2:40:59<2:34:41, 20.22s/it]
51%|█████▏ | 482/940 [2:41:17<2:29:44, 19.62s/it]
51%|█████▏ | 483/940 [2:41:36<2:27:03, 19.31s/it]
51%|█████▏ | 484/940 [2:41:56<2:28:08, 19.49s/it]
52%|█████▏ | 485/940 [2:42:13<2:23:20, 18.90s/it]
52%|█████▏ | 486/940 [2:42:31<2:20:56, 18.63s/it]
52%|█████▏ | 487/940 [2:42:52<2:25:05, 19.22s/it]
52%|█████▏ | 488/940 [2:43:14<2:30:49, 20.02s/it]
52%|█████▏ | 489/940 [2:43:34<2:30:22, 20.01s/it]
52%|█████▏ | 490/940 [2:43:53<2:28:05, 19.75s/it]
{'loss': '0.3572', 'grad_norm': '0.6864', 'learning_rate': '5.519e-06', 'epoch': '2.085'}
|
||
52%|█████▏ | 490/940 [2:43:53<2:28:05, 19.75s/it]
52%|█████▏ | 491/940 [2:44:13<2:28:29, 19.84s/it]
52%|█████▏ | 492/940 [2:44:32<2:26:36, 19.63s/it]
52%|█████▏ | 493/940 [2:44:52<2:25:56, 19.59s/it]
53%|█████▎ | 494/940 [2:45:11<2:25:38, 19.59s/it]
53%|█████▎ | 495/940 [2:45:31<2:26:40, 19.78s/it]
53%|█████▎ | 496/940 [2:45:49<2:22:35, 19.27s/it]
53%|█████▎ | 497/940 [2:46:10<2:26:09, 19.80s/it]
53%|█████▎ | 498/940 [2:46:30<2:26:09, 19.84s/it]
53%|█████▎ | 499/940 [2:46:53<2:31:48, 20.65s/it]
53%|█████▎ | 500/940 [2:47:12<2:28:04, 20.19s/it]
{'loss': '0.3511', 'grad_norm': '0.6616', 'learning_rate': '5.334e-06', 'epoch': '2.128'}
|
||
53%|█████▎ | 500/940 [2:47:12<2:28:04, 20.19s/it][INFO|trainer.py:3797] 2026-04-30 19:07:43,068 >> Saving model checkpoint to saves/qwen3_14b/coding/fft/checkpoint-500
|
||
[INFO|configuration_utils.py:432] 2026-04-30 19:07:43,078 >> Configuration saved in saves/qwen3_14b/coding/fft/checkpoint-500/config.json
|
||
[INFO|configuration_utils.py:803] 2026-04-30 19:07:43,082 >> Configuration saved in saves/qwen3_14b/coding/fft/checkpoint-500/generation_config.json
|
||
|
||
Writing model shards: 0%| | 0/1 [00:00<?, ?it/s][A
|
||
Writing model shards: 100%|██████████| 1/1 [00:45<00:00, 45.96s/it][A
Writing model shards: 100%|██████████| 1/1 [00:45<00:00, 45.96s/it]
|
||
[INFO|modeling_utils.py:3380] 2026-04-30 19:08:29,085 >> Model weights saved in saves/qwen3_14b/coding/fft/checkpoint-500/model.safetensors
|
||
[INFO|tokenization_utils_base.py:3224] 2026-04-30 19:08:29,092 >> chat template saved in saves/qwen3_14b/coding/fft/checkpoint-500/chat_template.jinja
|
||
[INFO|tokenization_utils_base.py:2078] 2026-04-30 19:08:29,098 >> tokenizer config file saved in saves/qwen3_14b/coding/fft/checkpoint-500/tokenizer_config.json
|
||
53%|█████▎ | 501/940 [2:48:41<4:58:02, 40.73s/it]
53%|█████▎ | 502/940 [2:48:59<4:08:00, 33.97s/it]
54%|█████▎ | 503/940 [2:49:19<3:37:45, 29.90s/it]
54%|█████▎ | 504/940 [2:49:38<3:13:12, 26.59s/it]
54%|█████▎ | 505/940 [2:49:59<3:00:56, 24.96s/it]
54%|█████▍ | 506/940 [2:50:20<2:50:17, 23.54s/it]
54%|█████▍ | 507/940 [2:50:43<2:49:02, 23.42s/it]
54%|█████▍ | 508/940 [2:51:03<2:40:52, 22.34s/it]
54%|█████▍ | 509/940 [2:51:23<2:36:42, 21.82s/it]
54%|█████▍ | 510/940 [2:51:41<2:28:01, 20.66s/it]
{'loss': '0.3635', 'grad_norm': '0.7258', 'learning_rate': '5.149e-06', 'epoch': '2.171'}
|
||
54%|█████▍ | 510/940 [2:51:41<2:28:01, 20.66s/it]
54%|█████▍ | 511/940 [2:52:00<2:22:56, 19.99s/it]
54%|█████▍ | 512/940 [2:52:23<2:29:32, 20.96s/it]
55%|█████▍ | 513/940 [2:52:43<2:28:31, 20.87s/it]
55%|█████▍ | 514/940 [2:53:04<2:28:19, 20.89s/it]
55%|█████▍ | 515/940 [2:53:26<2:30:12, 21.21s/it]
55%|█████▍ | 516/940 [2:53:48<2:30:58, 21.37s/it]
55%|█████▌ | 517/940 [2:54:08<2:27:26, 20.91s/it]
55%|█████▌ | 518/940 [2:54:27<2:23:52, 20.46s/it]
55%|█████▌ | 519/940 [2:54:45<2:17:47, 19.64s/it]
55%|█████▌ | 520/940 [2:55:05<2:18:30, 19.79s/it]
{'loss': '0.3607', 'grad_norm': '0.7153', 'learning_rate': '4.963e-06', 'epoch': '2.213'}
|
||
55%|█████▌ | 520/940 [2:55:05<2:18:30, 19.79s/it]
55%|█████▌ | 521/940 [2:55:24<2:16:21, 19.53s/it]
56%|█████▌ | 522/940 [2:55:43<2:14:08, 19.26s/it]
56%|█████▌ | 523/940 [2:56:00<2:10:42, 18.81s/it]
56%|█████▌ | 524/940 [2:56:21<2:13:34, 19.27s/it]
56%|█████▌ | 525/940 [2:56:40<2:13:19, 19.28s/it]
56%|█████▌ | 526/940 [2:57:02<2:17:34, 19.94s/it]
56%|█████▌ | 527/940 [2:57:21<2:16:35, 19.84s/it]
56%|█████▌ | 528/940 [2:57:39<2:12:13, 19.26s/it]
56%|█████▋ | 529/940 [2:58:00<2:15:40, 19.81s/it]
56%|█████▋ | 530/940 [2:58:20<2:14:38, 19.70s/it]
{'loss': '0.3447', 'grad_norm': '0.6809', 'learning_rate': '4.777e-06', 'epoch': '2.256'}
|
||
56%|█████▋ | 530/940 [2:58:20<2:14:38, 19.70s/it]
56%|█████▋ | 531/940 [2:58:42<2:20:18, 20.58s/it]
57%|█████▋ | 532/940 [2:59:01<2:16:19, 20.05s/it]
57%|█████▋ | 533/940 [2:59:22<2:16:47, 20.17s/it]
57%|█████▋ | 534/940 [2:59:41<2:14:44, 19.91s/it]
57%|█████▋ | 535/940 [3:00:01<2:14:40, 19.95s/it]
57%|█████▋ | 536/940 [3:00:19<2:11:34, 19.54s/it]
57%|█████▋ | 537/940 [3:00:38<2:09:09, 19.23s/it]
57%|█████▋ | 538/940 [3:00:58<2:10:46, 19.52s/it]
57%|█████▋ | 539/940 [3:01:17<2:08:25, 19.22s/it]
57%|█████▋ | 540/940 [3:01:37<2:10:27, 19.57s/it]
{'loss': '0.3704', 'grad_norm': '0.7152', 'learning_rate': '4.592e-06', 'epoch': '2.299'}
|
||
57%|█████▋ | 540/940 [3:01:37<2:10:27, 19.57s/it]
58%|█████▊ | 541/940 [3:01:55<2:07:36, 19.19s/it]
58%|█████▊ | 542/940 [3:02:14<2:06:15, 19.03s/it]
58%|█████▊ | 543/940 [3:02:33<2:05:21, 18.94s/it]
58%|█████▊ | 544/940 [3:02:55<2:11:35, 19.94s/it]
58%|█████▊ | 545/940 [3:03:13<2:08:08, 19.47s/it]
58%|█████▊ | 546/940 [3:03:34<2:10:21, 19.85s/it]
58%|█████▊ | 547/940 [3:03:54<2:10:26, 19.91s/it]
58%|█████▊ | 548/940 [3:04:13<2:07:49, 19.57s/it]
58%|█████▊ | 549/940 [3:04:31<2:05:25, 19.25s/it]
59%|█████▊ | 550/940 [3:04:51<2:05:15, 19.27s/it]
{'loss': '0.3504', 'grad_norm': '0.7301', 'learning_rate': '4.407e-06', 'epoch': '2.341'}
|
||
59%|█████▊ | 550/940 [3:04:51<2:05:15, 19.27s/it]
59%|█████▊ | 551/940 [3:05:09<2:02:45, 18.93s/it]
59%|█████▊ | 552/940 [3:05:27<2:00:11, 18.59s/it]
59%|█████▉ | 553/940 [3:05:46<2:00:21, 18.66s/it]
59%|█████▉ | 554/940 [3:06:06<2:02:46, 19.08s/it]
59%|█████▉ | 555/940 [3:06:25<2:02:09, 19.04s/it]
59%|█████▉ | 556/940 [3:06:43<2:00:45, 18.87s/it]
59%|█████▉ | 557/940 [3:07:02<1:59:58, 18.80s/it]
59%|█████▉ | 558/940 [3:07:24<2:07:06, 19.96s/it]
59%|█████▉ | 559/940 [3:07:44<2:07:00, 20.00s/it]
60%|█████▉ | 560/940 [3:08:04<2:05:56, 19.89s/it]
{'loss': '0.3553', 'grad_norm': '0.758', 'learning_rate': '4.223e-06', 'epoch': '2.384'}
|
||
60%|█████▉ | 560/940 [3:08:04<2:05:56, 19.89s/it]
60%|█████▉ | 561/940 [3:08:25<2:07:45, 20.23s/it]
60%|█████▉ | 562/940 [3:08:45<2:06:31, 20.08s/it]
60%|█████▉ | 563/940 [3:09:05<2:06:09, 20.08s/it]
60%|██████ | 564/940 [3:09:23<2:02:40, 19.58s/it]
60%|██████ | 565/940 [3:09:42<2:00:22, 19.26s/it]
60%|██████ | 566/940 [3:10:01<2:00:10, 19.28s/it]
60%|██████ | 567/940 [3:10:23<2:04:32, 20.03s/it]
60%|██████ | 568/940 [3:10:41<2:00:28, 19.43s/it]
61%|██████ | 569/940 [3:10:59<1:57:25, 18.99s/it]
61%|██████ | 570/940 [3:11:21<2:02:24, 19.85s/it]
{'loss': '0.3483', 'grad_norm': '0.6657', 'learning_rate': '4.04e-06', 'epoch': '2.427'}
|
||
61%|██████ | 570/940 [3:11:21<2:02:24, 19.85s/it]
61%|██████ | 571/940 [3:11:41<2:02:50, 19.98s/it]
61%|██████ | 572/940 [3:12:03<2:05:55, 20.53s/it]
61%|██████ | 573/940 [3:12:24<2:05:55, 20.59s/it]
61%|██████ | 574/940 [3:12:44<2:04:34, 20.42s/it]
61%|██████ | 575/940 [3:13:09<2:13:50, 22.00s/it]
61%|██████▏ | 576/940 [3:13:30<2:11:47, 21.72s/it]
61%|██████▏ | 577/940 [3:13:49<2:06:12, 20.86s/it]
61%|██████▏ | 578/940 [3:14:08<2:01:45, 20.18s/it]
62%|██████▏ | 579/940 [3:14:26<1:57:46, 19.57s/it]
62%|██████▏ | 580/940 [3:14:46<1:58:48, 19.80s/it]
{'loss': '0.3438', 'grad_norm': '0.6942', 'learning_rate': '3.859e-06', 'epoch': '2.469'}
|
||
62%|██████▏ | 580/940 [3:14:46<1:58:48, 19.80s/it]
62%|██████▏ | 581/940 [3:15:05<1:56:57, 19.55s/it]
62%|██████▏ | 582/940 [3:15:26<1:59:04, 19.96s/it]
62%|██████▏ | 583/940 [3:15:45<1:57:13, 19.70s/it]
62%|██████▏ | 584/940 [3:16:03<1:53:41, 19.16s/it]
62%|██████▏ | 585/940 [3:16:22<1:52:07, 18.95s/it]
62%|██████▏ | 586/940 [3:16:42<1:53:39, 19.26s/it]
62%|██████▏ | 587/940 [3:17:03<1:56:12, 19.75s/it]
63%|██████▎ | 588/940 [3:17:22<1:55:13, 19.64s/it]
63%|██████▎ | 589/940 [3:17:42<1:54:54, 19.64s/it]
63%|██████▎ | 590/940 [3:18:00<1:52:20, 19.26s/it]
{'loss': '0.3701', 'grad_norm': '0.7356', 'learning_rate': '3.679e-06', 'epoch': '2.512'}
|
||
63%|██████▎ | 590/940 [3:18:00<1:52:20, 19.26s/it]
63%|██████▎ | 591/940 [3:18:24<2:00:19, 20.69s/it]
63%|██████▎ | 592/940 [3:18:43<1:57:59, 20.34s/it]
63%|██████▎ | 593/940 [3:19:05<1:59:49, 20.72s/it]
63%|██████▎ | 594/940 [3:19:31<2:08:38, 22.31s/it]
63%|██████▎ | 595/940 [3:19:51<2:04:35, 21.67s/it]
63%|██████▎ | 596/940 [3:20:12<2:02:23, 21.35s/it]
64%|██████▎ | 597/940 [3:20:35<2:04:52, 21.84s/it]
64%|██████▎ | 598/940 [3:20:59<2:08:42, 22.58s/it]
64%|██████▎ | 599/940 [3:21:23<2:10:53, 23.03s/it]
64%|██████▍ | 600/940 [3:21:44<2:07:21, 22.48s/it]
{'loss': '0.3588', 'grad_norm': '0.7637', 'learning_rate': '3.501e-06', 'epoch': '2.555'}
|
||
64%|██████▍ | 600/940 [3:21:44<2:07:21, 22.48s/it]
64%|██████▍ | 601/940 [3:22:04<2:01:50, 21.57s/it]
64%|██████▍ | 602/940 [3:22:25<2:01:28, 21.56s/it]
64%|██████▍ | 603/940 [3:22:46<1:59:34, 21.29s/it]
64%|██████▍ | 604/940 [3:23:04<1:53:51, 20.33s/it]
64%|██████▍ | 605/940 [3:23:24<1:52:42, 20.19s/it]
64%|██████▍ | 606/940 [3:23:47<1:56:21, 20.90s/it]
65%|██████▍ | 607/940 [3:24:05<1:52:03, 20.19s/it]
65%|██████▍ | 608/940 [3:24:25<1:51:26, 20.14s/it]
65%|██████▍ | 609/940 [3:24:45<1:51:04, 20.13s/it]
65%|██████▍ | 610/940 [3:25:06<1:51:51, 20.34s/it]
{'loss': '0.3827', 'grad_norm': '0.7925', 'learning_rate': '3.325e-06', 'epoch': '2.597'}
|
||
65%|██████▍ | 610/940 [3:25:06<1:51:51, 20.34s/it]
65%|██████▌ | 611/940 [3:25:25<1:48:45, 19.83s/it]
65%|██████▌ | 612/940 [3:25:44<1:46:49, 19.54s/it]
65%|██████▌ | 613/940 [3:26:04<1:47:14, 19.68s/it]
65%|██████▌ | 614/940 [3:26:23<1:46:02, 19.52s/it]
65%|██████▌ | 615/940 [3:26:43<1:46:28, 19.66s/it]
66%|██████▌ | 616/940 [3:27:02<1:45:21, 19.51s/it]
66%|██████▌ | 617/940 [3:27:22<1:46:49, 19.84s/it]
66%|██████▌ | 618/940 [3:27:43<1:47:35, 20.05s/it]
66%|██████▌ | 619/940 [3:28:05<1:50:33, 20.66s/it]
66%|██████▌ | 620/940 [3:28:24<1:47:03, 20.07s/it]
{'loss': '0.3599', 'grad_norm': '0.7416', 'learning_rate': '3.151e-06', 'epoch': '2.64'}
|
||
66%|██████▌ | 620/940 [3:28:24<1:47:03, 20.07s/it]
66%|██████▌ | 621/940 [3:28:42<1:43:58, 19.56s/it]
66%|██████▌ | 622/940 [3:29:00<1:40:35, 18.98s/it]
66%|██████▋ | 623/940 [3:29:21<1:43:59, 19.68s/it]
66%|██████▋ | 624/940 [3:29:42<1:45:49, 20.09s/it]
66%|██████▋ | 625/940 [3:30:04<1:47:35, 20.49s/it]
67%|██████▋ | 626/940 [3:30:23<1:45:04, 20.08s/it]
67%|██████▋ | 627/940 [3:30:44<1:45:56, 20.31s/it]
67%|██████▋ | 628/940 [3:31:02<1:42:39, 19.74s/it]
67%|██████▋ | 629/940 [3:31:20<1:40:11, 19.33s/it]
67%|██████▋ | 630/940 [3:31:39<1:38:47, 19.12s/it]
{'loss': '0.3487', 'grad_norm': '0.7405', 'learning_rate': '2.98e-06', 'epoch': '2.683'}
|
||
67%|██████▋ | 630/940 [3:31:39<1:38:47, 19.12s/it]
67%|██████▋ | 631/940 [3:31:58<1:38:14, 19.07s/it]
67%|██████▋ | 632/940 [3:32:17<1:38:11, 19.13s/it]
67%|██████▋ | 633/940 [3:32:39<1:41:13, 19.78s/it]
67%|██████▋ | 634/940 [3:32:58<1:40:19, 19.67s/it]
68%|██████▊ | 635/940 [3:33:18<1:40:49, 19.83s/it]
68%|██████▊ | 636/940 [3:33:37<1:38:43, 19.49s/it]
68%|██████▊ | 637/940 [3:33:58<1:40:46, 19.96s/it]
68%|██████▊ | 638/940 [3:34:20<1:44:18, 20.72s/it]
68%|██████▊ | 639/940 [3:34:41<1:43:18, 20.59s/it]
68%|██████▊ | 640/940 [3:35:03<1:45:15, 21.05s/it]
{'loss': '0.3588', 'grad_norm': '0.752', 'learning_rate': '2.811e-06', 'epoch': '2.725'}
|
||
68%|██████▊ | 640/940 [3:35:03<1:45:15, 21.05s/it]
68%|██████▊ | 641/940 [3:35:23<1:43:43, 20.81s/it]
68%|██████▊ | 642/940 [3:35:45<1:45:27, 21.23s/it]
68%|██████▊ | 643/940 [3:36:07<1:45:56, 21.40s/it]
69%|██████▊ | 644/940 [3:36:26<1:42:15, 20.73s/it]
69%|██████▊ | 645/940 [3:36:46<1:40:15, 20.39s/it]
69%|██████▊ | 646/940 [3:37:05<1:37:49, 19.96s/it]
69%|██████▉ | 647/940 [3:37:26<1:39:39, 20.41s/it]
69%|██████▉ | 648/940 [3:37:46<1:38:53, 20.32s/it]
69%|██████▉ | 649/940 [3:38:07<1:38:30, 20.31s/it]
69%|██████▉ | 650/940 [3:38:28<1:39:00, 20.48s/it]
{'loss': '0.3493', 'grad_norm': '0.7427', 'learning_rate': '2.646e-06', 'epoch': '2.768'}
|
||
69%|██████▉ | 650/940 [3:38:28<1:39:00, 20.48s/it]
69%|██████▉ | 651/940 [3:38:45<1:34:40, 19.66s/it]
69%|██████▉ | 652/940 [3:39:04<1:33:02, 19.38s/it]
69%|██████▉ | 653/940 [3:39:26<1:36:05, 20.09s/it]
70%|██████▉ | 654/940 [3:39:45<1:35:00, 19.93s/it]
70%|██████▉ | 655/940 [3:40:03<1:31:43, 19.31s/it]
70%|██████▉ | 656/940 [3:40:23<1:31:38, 19.36s/it]
70%|██████▉ | 657/940 [3:40:40<1:29:11, 18.91s/it]
70%|███████ | 658/940 [3:41:00<1:29:32, 19.05s/it]
70%|███████ | 659/940 [3:41:21<1:32:43, 19.80s/it]
70%|███████ | 660/940 [3:41:45<1:38:14, 21.05s/it]
{'loss': '0.3375', 'grad_norm': '0.7397', 'learning_rate': '2.484e-06', 'epoch': '2.811'}
|
||
70%|███████ | 660/940 [3:41:45<1:38:14, 21.05s/it]
70%|███████ | 661/940 [3:42:04<1:34:52, 20.40s/it]
70%|███████ | 662/940 [3:42:25<1:34:21, 20.36s/it]
71%|███████ | 663/940 [3:42:45<1:33:51, 20.33s/it]
71%|███████ | 664/940 [3:43:05<1:33:31, 20.33s/it]
71%|███████ | 665/940 [3:43:25<1:32:47, 20.25s/it]
71%|███████ | 666/940 [3:43:43<1:29:30, 19.60s/it]
71%|███████ | 667/940 [3:44:05<1:32:08, 20.25s/it]
71%|███████ | 668/940 [3:44:29<1:37:00, 21.40s/it]
71%|███████ | 669/940 [3:44:49<1:34:58, 21.03s/it]
71%|███████▏ | 670/940 [3:45:10<1:33:40, 20.82s/it]
{'loss': '0.3639', 'grad_norm': '0.7484', 'learning_rate': '2.325e-06', 'epoch': '2.853'}
|
||
71%|███████▏ | 670/940 [3:45:10<1:33:40, 20.82s/it]
71%|███████▏ | 671/940 [3:45:31<1:33:42, 20.90s/it]
71%|███████▏ | 672/940 [3:45:51<1:32:26, 20.70s/it]
72%|███████▏ | 673/940 [3:46:11<1:31:16, 20.51s/it]
72%|███████▏ | 674/940 [3:46:29<1:27:09, 19.66s/it]
72%|███████▏ | 675/940 [3:46:47<1:25:00, 19.25s/it]
72%|███████▏ | 676/940 [3:47:05<1:23:23, 18.95s/it]
72%|███████▏ | 677/940 [3:47:24<1:22:13, 18.76s/it]
72%|███████▏ | 678/940 [3:47:44<1:23:41, 19.17s/it]
72%|███████▏ | 679/940 [3:48:06<1:27:29, 20.11s/it]
72%|███████▏ | 680/940 [3:48:29<1:31:15, 21.06s/it]
{'loss': '0.3687', 'grad_norm': '0.7073', 'learning_rate': '2.17e-06', 'epoch': '2.896'}
|
||
72%|███████▏ | 680/940 [3:48:29<1:31:15, 21.06s/it]
72%|███████▏ | 681/940 [3:48:49<1:29:40, 20.78s/it]
73%|███████▎ | 682/940 [3:49:08<1:26:35, 20.14s/it]
73%|███████▎ | 683/940 [3:49:27<1:24:55, 19.83s/it]
73%|███████▎ | 684/940 [3:49:49<1:26:47, 20.34s/it]
73%|███████▎ | 685/940 [3:50:08<1:25:47, 20.19s/it]
73%|███████▎ | 686/940 [3:50:26<1:22:33, 19.50s/it]
73%|███████▎ | 687/940 [3:50:48<1:24:57, 20.15s/it]
73%|███████▎ | 688/940 [3:51:08<1:24:22, 20.09s/it]
73%|███████▎ | 689/940 [3:51:27<1:22:39, 19.76s/it]
73%|███████▎ | 690/940 [3:51:47<1:22:32, 19.81s/it]
{'loss': '0.3675', 'grad_norm': '0.7069', 'learning_rate': '2.019e-06', 'epoch': '2.939'}
|
||
73%|███████▎ | 690/940 [3:51:47<1:22:32, 19.81s/it]
74%|███████▎ | 691/940 [3:52:07<1:22:28, 19.87s/it]
74%|███████▎ | 692/940 [3:52:28<1:23:08, 20.12s/it]
74%|███████▎ | 693/940 [3:52:46<1:20:15, 19.50s/it]
74%|███████▍ | 694/940 [3:53:05<1:19:11, 19.32s/it]
74%|███████▍ | 695/940 [3:53:25<1:20:30, 19.71s/it]
74%|███████▍ | 696/940 [3:53:48<1:24:10, 20.70s/it]
74%|███████▍ | 697/940 [3:54:07<1:22:06, 20.27s/it]
74%|███████▍ | 698/940 [3:54:29<1:23:41, 20.75s/it]
74%|███████▍ | 699/940 [3:54:48<1:20:29, 20.04s/it]
74%|███████▍ | 700/940 [3:55:06<1:18:00, 19.50s/it]
{'loss': '0.3496', 'grad_norm': '0.7173', 'learning_rate': '1.872e-06', 'epoch': '2.981'}
|
||
74%|███████▍ | 700/940 [3:55:06<1:18:00, 19.50s/it]
75%|███████▍ | 701/940 [3:55:25<1:17:20, 19.42s/it]
75%|███████▍ | 702/940 [3:55:43<1:15:27, 19.02s/it]
75%|███████▍ | 703/940 [3:56:04<1:16:44, 19.43s/it]
75%|███████▍ | 704/940 [3:56:22<1:15:29, 19.19s/it]
75%|███████▌ | 705/940 [3:56:29<1:00:40, 15.49s/it]
75%|███████▌ | 706/940 [3:56:51<1:07:40, 17.35s/it]
75%|███████▌ | 707/940 [3:57:11<1:10:08, 18.06s/it]
75%|███████▌ | 708/940 [3:57:30<1:11:50, 18.58s/it]
75%|███████▌ | 709/940 [3:57:54<1:16:58, 19.99s/it]
76%|███████▌ | 710/940 [3:58:12<1:15:03, 19.58s/it]
{'loss': '0.3147', 'grad_norm': '0.8301', 'learning_rate': '1.73e-06', 'epoch': '3.021'}
|
||
76%|███████▌ | 710/940 [3:58:12<1:15:03, 19.58s/it]
76%|███████▌ | 711/940 [3:58:31<1:13:22, 19.23s/it]
76%|███████▌ | 712/940 [3:58:51<1:14:25, 19.58s/it]
76%|███████▌ | 713/940 [3:59:10<1:13:22, 19.39s/it]
76%|███████▌ | 714/940 [3:59:28<1:11:45, 19.05s/it]
76%|███████▌ | 715/940 [3:59:46<1:10:24, 18.78s/it]
76%|███████▌ | 716/940 [4:00:08<1:13:08, 19.59s/it]
76%|███████▋ | 717/940 [4:00:29<1:14:00, 19.91s/it]
76%|███████▋ | 718/940 [4:00:48<1:13:23, 19.83s/it]
76%|███████▋ | 719/940 [4:01:08<1:12:48, 19.77s/it]
77%|███████▋ | 720/940 [4:01:28<1:12:37, 19.81s/it]
{'loss': '0.2616', 'grad_norm': '1.07', 'learning_rate': '1.591e-06', 'epoch': '3.064'}
|
||
77%|███████▋ | 720/940 [4:01:28<1:12:37, 19.81s/it]
77%|███████▋ | 721/940 [4:01:46<1:11:00, 19.45s/it]
77%|███████▋ | 722/940 [4:02:05<1:10:06, 19.30s/it]
77%|███████▋ | 723/940 [4:02:28<1:12:59, 20.18s/it]
77%|███████▋ | 724/940 [4:02:48<1:12:50, 20.23s/it]
77%|███████▋ | 725/940 [4:03:09<1:13:27, 20.50s/it]
77%|███████▋ | 726/940 [4:03:27<1:10:46, 19.84s/it]
77%|███████▋ | 727/940 [4:03:49<1:12:30, 20.42s/it]
77%|███████▋ | 728/940 [4:04:08<1:10:06, 19.84s/it]
78%|███████▊ | 729/940 [4:04:27<1:08:58, 19.61s/it]
78%|███████▊ | 730/940 [4:04:48<1:10:01, 20.01s/it]
{'loss': '0.2639', 'grad_norm': '0.802', 'learning_rate': '1.458e-06', 'epoch': '3.107'}
|
||
78%|███████▊ | 730/940 [4:04:48<1:10:01, 20.01s/it]
78%|███████▊ | 731/940 [4:05:06<1:08:00, 19.52s/it]
78%|███████▊ | 732/940 [4:05:24<1:05:41, 18.95s/it]
78%|███████▊ | 733/940 [4:05:44<1:06:28, 19.27s/it]
78%|███████▊ | 734/940 [4:06:05<1:08:09, 19.85s/it]
78%|███████▊ | 735/940 [4:06:25<1:07:55, 19.88s/it]
78%|███████▊ | 736/940 [4:06:43<1:05:52, 19.38s/it]
78%|███████▊ | 737/940 [4:07:04<1:06:48, 19.75s/it]
79%|███████▊ | 738/940 [4:07:24<1:06:59, 19.90s/it]
79%|███████▊ | 739/940 [4:07:43<1:06:01, 19.71s/it]
79%|███████▊ | 740/940 [4:08:04<1:06:47, 20.04s/it]
{'loss': '0.2392', 'grad_norm': '0.8001', 'learning_rate': '1.329e-06', 'epoch': '3.149'}
|
||
79%|███████▊ | 740/940 [4:08:04<1:06:47, 20.04s/it]
79%|███████▉ | 741/940 [4:08:23<1:05:48, 19.84s/it]
79%|███████▉ | 742/940 [4:08:43<1:05:12, 19.76s/it]
79%|███████▉ | 743/940 [4:09:04<1:05:50, 20.05s/it]
79%|███████▉ | 744/940 [4:09:25<1:06:36, 20.39s/it]
79%|███████▉ | 745/940 [4:09:47<1:08:05, 20.95s/it]
79%|███████▉ | 746/940 [4:10:05<1:05:05, 20.13s/it]
79%|███████▉ | 747/940 [4:10:26<1:05:06, 20.24s/it]
80%|███████▉ | 748/940 [4:10:45<1:03:30, 19.84s/it]
80%|███████▉ | 749/940 [4:11:05<1:03:19, 19.89s/it]
80%|███████▉ | 750/940 [4:11:24<1:02:54, 19.86s/it]
{'loss': '0.2672', 'grad_norm': '0.8249', 'learning_rate': '1.206e-06', 'epoch': '3.192'}
|
||
80%|███████▉ | 750/940 [4:11:24<1:02:54, 19.86s/it]
80%|███████▉ | 751/940 [4:11:42<1:00:42, 19.27s/it]
80%|████████ | 752/940 [4:12:02<1:01:05, 19.50s/it]
80%|████████ | 753/940 [4:12:24<1:02:44, 20.13s/it]
80%|████████ | 754/940 [4:12:43<1:01:34, 19.86s/it]
80%|████████ | 755/940 [4:13:02<59:58, 19.45s/it]
80%|████████ | 756/940 [4:13:22<1:00:06, 19.60s/it]
81%|████████ | 757/940 [4:13:45<1:03:03, 20.67s/it]
81%|████████ | 758/940 [4:14:04<1:01:12, 20.18s/it]
81%|████████ | 759/940 [4:14:22<59:07, 19.60s/it]
81%|████████ | 760/940 [4:14:42<58:45, 19.59s/it]
{'loss': '0.2474', 'grad_norm': '0.9324', 'learning_rate': '1.088e-06', 'epoch': '3.235'}
|
||
81%|████████ | 760/940 [4:14:42<58:45, 19.59s/it]
81%|████████ | 761/940 [4:15:02<58:46, 19.70s/it]
81%|████████ | 762/940 [4:15:23<59:35, 20.08s/it]
81%|████████ | 763/940 [4:15:43<59:04, 20.02s/it]
81%|████████▏ | 764/940 [4:16:01<57:28, 19.60s/it]
81%|████████▏ | 765/940 [4:16:20<56:20, 19.32s/it]
81%|████████▏ | 766/940 [4:16:42<58:15, 20.09s/it]
82%|████████▏ | 767/940 [4:17:05<1:00:51, 21.11s/it]
82%|████████▏ | 768/940 [4:17:25<59:39, 20.81s/it]
82%|████████▏ | 769/940 [4:17:45<58:06, 20.39s/it]
82%|████████▏ | 770/940 [4:18:06<58:10, 20.53s/it]
{'loss': '0.252', 'grad_norm': '0.798', 'learning_rate': '9.746e-07', 'epoch': '3.277'}
|
||
82%|████████▏ | 770/940 [4:18:06<58:10, 20.53s/it]
82%|████████▏ | 771/940 [4:18:24<56:05, 19.91s/it]
82%|████████▏ | 772/940 [4:18:44<55:44, 19.91s/it]
82%|████████▏ | 773/940 [4:19:04<55:55, 20.09s/it]
82%|████████▏ | 774/940 [4:19:24<54:46, 19.80s/it]
82%|████████▏ | 775/940 [4:19:43<54:03, 19.66s/it]
83%|████████▎ | 776/940 [4:20:02<53:13, 19.47s/it]
83%|████████▎ | 777/940 [4:20:22<53:29, 19.69s/it]
83%|████████▎ | 778/940 [4:20:44<54:55, 20.34s/it]
83%|████████▎ | 779/940 [4:21:05<55:26, 20.66s/it]
83%|████████▎ | 780/940 [4:21:25<54:12, 20.33s/it]
{'loss': '0.2524', 'grad_norm': '0.8369', 'learning_rate': '8.673e-07', 'epoch': '3.32'}
|
||
83%|████████▎ | 780/940 [4:21:25<54:12, 20.33s/it]
83%|████████▎ | 781/940 [4:21:43<52:06, 19.67s/it]
83%|████████▎ | 782/940 [4:22:02<51:11, 19.44s/it]
83%|████████▎ | 783/940 [4:22:20<50:06, 19.15s/it]
83%|████████▎ | 784/940 [4:22:39<49:08, 18.90s/it]
84%|████████▎ | 785/940 [4:22:58<48:46, 18.88s/it]
84%|████████▎ | 786/940 [4:23:17<48:41, 18.97s/it]
84%|████████▎ | 787/940 [4:23:43<53:46, 21.09s/it]
84%|████████▍ | 788/940 [4:24:01<51:27, 20.31s/it]
84%|████████▍ | 789/940 [4:24:22<51:21, 20.41s/it]
84%|████████▍ | 790/940 [4:24:41<50:18, 20.12s/it]
{'loss': '0.26', 'grad_norm': '0.7617', 'learning_rate': '7.657e-07', 'epoch': '3.363'}
|
||
84%|████████▍ | 790/940 [4:24:41<50:18, 20.12s/it]
84%|████████▍ | 791/940 [4:25:02<50:37, 20.38s/it]
84%|████████▍ | 792/940 [4:25:24<51:02, 20.69s/it]
84%|████████▍ | 793/940 [4:25:42<48:52, 19.95s/it]
84%|████████▍ | 794/940 [4:26:01<47:32, 19.54s/it]
85%|████████▍ | 795/940 [4:26:23<49:38, 20.54s/it]
85%|████████▍ | 796/940 [4:26:45<50:19, 20.97s/it]
85%|████████▍ | 797/940 [4:27:07<50:24, 21.15s/it]
85%|████████▍ | 798/940 [4:27:25<47:46, 20.19s/it]
85%|████████▌ | 799/940 [4:27:46<48:20, 20.57s/it]
85%|████████▌ | 800/940 [4:28:09<49:18, 21.13s/it]
{'loss': '0.2527', 'grad_norm': '0.8667', 'learning_rate': '6.699e-07', 'epoch': '3.405'}
|
||
85%|████████▌ | 800/940 [4:28:09<49:18, 21.13s/it]
85%|████████▌ | 801/940 [4:28:27<46:46, 20.19s/it]
85%|████████▌ | 802/940 [4:28:46<45:37, 19.84s/it]
85%|████████▌ | 803/940 [4:29:10<48:19, 21.17s/it]
86%|████████▌ | 804/940 [4:29:29<46:05, 20.34s/it]
86%|████████▌ | 805/940 [4:29:49<45:43, 20.32s/it]
86%|████████▌ | 806/940 [4:30:08<44:30, 19.93s/it]
86%|████████▌ | 807/940 [4:30:27<43:26, 19.60s/it]
86%|████████▌ | 808/940 [4:30:45<42:27, 19.30s/it]
86%|████████▌ | 809/940 [4:31:06<43:05, 19.74s/it]
86%|████████▌ | 810/940 [4:31:32<47:06, 21.74s/it]
{'loss': '0.2648', 'grad_norm': '0.7931', 'learning_rate': '5.8e-07', 'epoch': '3.448'}
|
||
86%|████████▌ | 810/940 [4:31:32<47:06, 21.74s/it]
86%|████████▋ | 811/940 [4:32:00<50:25, 23.45s/it]
86%|████████▋ | 812/940 [4:32:24<50:12, 23.53s/it]
86%|████████▋ | 813/940 [4:32:43<47:21, 22.37s/it]
87%|████████▋ | 814/940 [4:33:04<46:04, 21.94s/it]
87%|████████▋ | 815/940 [4:33:24<44:06, 21.17s/it]
87%|████████▋ | 816/940 [4:33:44<43:08, 20.87s/it]
87%|████████▋ | 817/940 [4:34:02<41:02, 20.02s/it]
87%|████████▋ | 818/940 [4:34:24<41:55, 20.62s/it]
87%|████████▋ | 819/940 [4:34:42<39:57, 19.82s/it]
87%|████████▋ | 820/940 [4:35:01<39:32, 19.77s/it]
{'loss': '0.2601', 'grad_norm': '0.918', 'learning_rate': '4.963e-07', 'epoch': '3.491'}
|
||
87%|████████▋ | 820/940 [4:35:01<39:32, 19.77s/it]
87%|████████▋ | 821/940 [4:35:21<38:54, 19.62s/it]
87%|████████▋ | 822/940 [4:35:41<38:47, 19.72s/it]
88%|████████▊ | 823/940 [4:35:59<37:49, 19.40s/it]
88%|████████▊ | 824/940 [4:36:22<39:08, 20.25s/it]
88%|████████▊ | 825/940 [4:36:40<37:44, 19.70s/it]
88%|████████▊ | 826/940 [4:36:58<36:30, 19.21s/it]
88%|████████▊ | 827/940 [4:37:16<35:41, 18.95s/it]
88%|████████▊ | 828/940 [4:37:35<35:23, 18.96s/it]
88%|████████▊ | 829/940 [4:37:54<34:40, 18.75s/it]
88%|████████▊ | 830/940 [4:38:11<33:52, 18.48s/it]
{'loss': '0.2649', 'grad_norm': '0.9272', 'learning_rate': '4.188e-07', 'epoch': '3.533'}
|
||
88%|████████▊ | 830/940 [4:38:11<33:52, 18.48s/it]
88%|████████▊ | 831/940 [4:38:32<34:44, 19.13s/it]
89%|████████▊ | 832/940 [4:38:51<34:30, 19.17s/it]
89%|████████▊ | 833/940 [4:39:10<33:52, 19.00s/it]
89%|████████▊ | 834/940 [4:39:32<35:04, 19.86s/it]
89%|████████▉ | 835/940 [4:39:52<34:43, 19.84s/it]
89%|████████▉ | 836/940 [4:40:12<34:53, 20.13s/it]
89%|████████▉ | 837/940 [4:40:30<33:26, 19.48s/it]
89%|████████▉ | 838/940 [4:40:48<32:15, 18.98s/it]
89%|████████▉ | 839/940 [4:41:08<32:26, 19.27s/it]
89%|████████▉ | 840/940 [4:41:30<33:16, 19.97s/it]
{'loss': '0.2623', 'grad_norm': '0.8183', 'learning_rate': '3.476e-07', 'epoch': '3.576'}
|
||
89%|████████▉ | 840/940 [4:41:30<33:16, 19.97s/it]
89%|████████▉ | 841/940 [4:41:53<34:26, 20.88s/it]
90%|████████▉ | 842/940 [4:42:11<32:38, 19.98s/it]
90%|████████▉ | 843/940 [4:42:30<31:59, 19.79s/it]
90%|████████▉ | 844/940 [4:42:48<30:44, 19.21s/it]
90%|████████▉ | 845/940 [4:43:07<30:14, 19.10s/it]
90%|█████████ | 846/940 [4:43:26<29:57, 19.12s/it]
90%|█████████ | 847/940 [4:43:47<30:47, 19.86s/it]
90%|█████████ | 848/940 [4:44:09<31:13, 20.37s/it]
90%|█████████ | 849/940 [4:44:31<31:43, 20.91s/it]
90%|█████████ | 850/940 [4:44:53<31:45, 21.17s/it]
{'loss': '0.2524', 'grad_norm': '0.794', 'learning_rate': '2.828e-07', 'epoch': '3.619'}
|
||
90%|█████████ | 850/940 [4:44:53<31:45, 21.17s/it]
91%|█████████ | 851/940 [4:45:16<32:11, 21.71s/it]
91%|█████████ | 852/940 [4:45:35<30:34, 20.85s/it]
91%|█████████ | 853/940 [4:45:52<28:44, 19.83s/it]
91%|█████████ | 854/940 [4:46:10<27:37, 19.28s/it]
91%|█████████ | 855/940 [4:46:31<27:58, 19.74s/it]
91%|█████████ | 856/940 [4:46:51<27:51, 19.90s/it]
91%|█████████ | 857/940 [4:47:11<27:39, 20.00s/it]
91%|█████████▏| 858/940 [4:47:32<27:22, 20.03s/it]
91%|█████████▏| 859/940 [4:47:52<27:14, 20.18s/it]
91%|█████████▏| 860/940 [4:48:13<27:11, 20.39s/it]
{'loss': '0.2619', 'grad_norm': '0.901', 'learning_rate': '2.245e-07', 'epoch': '3.661'}
|
||
91%|█████████▏| 860/940 [4:48:13<27:11, 20.39s/it]
92%|█████████▏| 861/940 [4:48:34<27:07, 20.60s/it]
92%|█████████▏| 862/940 [4:48:55<26:59, 20.76s/it]
92%|█████████▏| 863/940 [4:49:16<26:36, 20.73s/it]
92%|█████████▏| 864/940 [4:49:36<26:06, 20.61s/it]
92%|█████████▏| 865/940 [4:49:57<25:40, 20.54s/it]
92%|█████████▏| 866/940 [4:50:17<25:14, 20.47s/it]
92%|█████████▏| 867/940 [4:50:38<25:01, 20.57s/it]
92%|█████████▏| 868/940 [4:50:59<25:01, 20.86s/it]
92%|█████████▏| 869/940 [4:51:21<24:57, 21.10s/it]
93%|█████████▎| 870/940 [4:51:42<24:41, 21.16s/it]
{'loss': '0.2524', 'grad_norm': '0.8711', 'learning_rate': '1.728e-07', 'epoch': '3.704'}
|
||
93%|█████████▎| 870/940 [4:51:42<24:41, 21.16s/it]
93%|█████████▎| 871/940 [4:52:03<24:08, 21.00s/it]
93%|█████████▎| 872/940 [4:52:25<24:13, 21.37s/it]
93%|█████████▎| 873/940 [4:52:48<24:23, 21.85s/it]
93%|█████████▎| 874/940 [4:53:08<23:33, 21.42s/it]
93%|█████████▎| 875/940 [4:53:30<23:22, 21.57s/it]
93%|█████████▎| 876/940 [4:53:52<23:07, 21.69s/it]
93%|█████████▎| 877/940 [4:54:13<22:18, 21.24s/it]
93%|█████████▎| 878/940 [4:54:34<22:05, 21.37s/it]
94%|█████████▎| 879/940 [4:54:57<22:09, 21.80s/it]
94%|█████████▎| 880/940 [4:55:17<21:09, 21.15s/it]
{'loss': '0.2483', 'grad_norm': '0.8006', 'learning_rate': '1.277e-07', 'epoch': '3.747'}
|
||
94%|█████████▎| 880/940 [4:55:17<21:09, 21.15s/it]
94%|█████████▎| 881/940 [4:55:35<19:52, 20.21s/it]
94%|█████████▍| 882/940 [4:55:54<19:23, 20.07s/it]
94%|█████████▍| 883/940 [4:56:13<18:32, 19.51s/it]
94%|█████████▍| 884/940 [4:56:35<18:55, 20.28s/it]
94%|█████████▍| 885/940 [4:56:53<17:55, 19.56s/it]
94%|█████████▍| 886/940 [4:57:13<17:57, 19.95s/it]
94%|█████████▍| 887/940 [4:57:33<17:38, 19.97s/it]
94%|█████████▍| 888/940 [4:57:57<18:17, 21.11s/it]
95%|█████████▍| 889/940 [4:58:16<17:25, 20.50s/it]
95%|█████████▍| 890/940 [4:58:34<16:28, 19.78s/it]
{'loss': '0.2666', 'grad_norm': '0.8466', 'learning_rate': '8.94e-08', 'epoch': '3.789'}
|
||
95%|█████████▍| 890/940 [4:58:34<16:28, 19.78s/it]
95%|█████████▍| 891/940 [4:58:53<15:50, 19.40s/it]
95%|█████████▍| 892/940 [4:59:15<16:04, 20.10s/it]
95%|█████████▌| 893/940 [4:59:36<16:00, 20.43s/it]
95%|█████████▌| 894/940 [4:59:58<16:06, 21.01s/it]
95%|█████████▌| 895/940 [5:00:20<15:56, 21.25s/it]
95%|█████████▌| 896/940 [5:00:41<15:26, 21.05s/it]
95%|█████████▌| 897/940 [5:01:01<14:59, 20.91s/it]
96%|█████████▌| 898/940 [5:01:23<14:51, 21.22s/it]
96%|█████████▌| 899/940 [5:01:46<14:45, 21.60s/it]
96%|█████████▌| 900/940 [5:02:08<14:37, 21.95s/it]
{'loss': '0.2557', 'grad_norm': '0.8011', 'learning_rate': '5.784e-08', 'epoch': '3.832'}
|
||
96%|█████████▌| 900/940 [5:02:08<14:37, 21.95s/it]
96%|█████████▌| 901/940 [5:02:32<14:35, 22.44s/it]
96%|█████████▌| 902/940 [5:02:55<14:24, 22.74s/it]
96%|█████████▌| 903/940 [5:03:16<13:42, 22.23s/it]
96%|█████████▌| 904/940 [5:03:37<13:00, 21.67s/it]
96%|█████████▋| 905/940 [5:04:01<13:09, 22.55s/it]
96%|█████████▋| 906/940 [5:04:22<12:29, 22.05s/it]
96%|█████████▋| 907/940 [5:04:44<12:05, 21.99s/it]
97%|█████████▋| 908/940 [5:05:02<11:08, 20.88s/it]
97%|█████████▋| 909/940 [5:05:21<10:24, 20.16s/it]
97%|█████████▋| 910/940 [5:05:42<10:15, 20.52s/it]
{'loss': '0.2542', 'grad_norm': '0.7995', 'learning_rate': '3.309e-08', 'epoch': '3.875'}
|
||
97%|█████████▋| 910/940 [5:05:42<10:15, 20.52s/it]
97%|█████████▋| 911/940 [5:06:04<10:04, 20.85s/it]
97%|█████████▋| 912/940 [5:06:25<09:49, 21.07s/it]
97%|█████████▋| 913/940 [5:06:46<09:26, 20.98s/it]
97%|█████████▋| 914/940 [5:07:06<08:53, 20.51s/it]
97%|█████████▋| 915/940 [5:07:26<08:29, 20.39s/it]
97%|█████████▋| 916/940 [5:07:46<08:10, 20.43s/it]
98%|█████████▊| 917/940 [5:08:05<07:35, 19.79s/it]
98%|█████████▊| 918/940 [5:08:24<07:15, 19.78s/it]
98%|█████████▊| 919/940 [5:08:42<06:43, 19.20s/it]
98%|█████████▊| 920/940 [5:09:02<06:28, 19.40s/it]
{'loss': '0.2418', 'grad_norm': '0.8056', 'learning_rate': '1.52e-08', 'epoch': '3.917'}
|
||
98%|█████████▊| 920/940 [5:09:02<06:28, 19.40s/it]
98%|█████████▊| 921/940 [5:09:21<06:07, 19.32s/it]
98%|█████████▊| 922/940 [5:09:39<05:41, 18.98s/it]
98%|█████████▊| 923/940 [5:09:57<05:17, 18.70s/it]
98%|█████████▊| 924/940 [5:10:19<05:12, 19.53s/it]
98%|█████████▊| 925/940 [5:10:37<04:46, 19.10s/it]
99%|█████████▊| 926/940 [5:10:57<04:30, 19.31s/it]
99%|█████████▊| 927/940 [5:11:15<04:08, 19.11s/it]
99%|█████████▊| 928/940 [5:11:35<03:50, 19.23s/it]
99%|█████████▉| 929/940 [5:11:54<03:32, 19.32s/it]
99%|█████████▉| 930/940 [5:12:13<03:10, 19.07s/it]
{'loss': '0.2599', 'grad_norm': '0.8602', 'learning_rate': '4.171e-09', 'epoch': '3.96'}
|
||
99%|█████████▉| 930/940 [5:12:13<03:10, 19.07s/it]
99%|█████████▉| 931/940 [5:12:31<02:48, 18.68s/it]
99%|█████████▉| 932/940 [5:12:49<02:27, 18.46s/it]
99%|█████████▉| 933/940 [5:13:09<02:12, 18.96s/it]
99%|█████████▉| 934/940 [5:13:27<01:52, 18.70s/it]
99%|█████████▉| 935/940 [5:13:45<01:33, 18.66s/it]
100%|█████████▉| 936/940 [5:14:08<01:19, 19.78s/it]
100%|█████████▉| 937/940 [5:14:26<00:57, 19.29s/it]
100%|█████████▉| 938/940 [5:14:52<00:42, 21.21s/it]
100%|█████████▉| 939/940 [5:15:11<00:20, 20.51s/it]
100%|██████████| 940/940 [5:15:18<00:00, 16.47s/it]
{'loss': '0.2301', 'grad_norm': '1.38', 'learning_rate': '3.447e-11', 'epoch': '4'}
|
||
100%|██████████| 940/940 [5:15:18<00:00, 16.47s/it][INFO|trainer.py:3797] 2026-04-30 21:35:49,247 >> Saving model checkpoint to saves/qwen3_14b/coding/fft/checkpoint-940
|
||
[INFO|configuration_utils.py:432] 2026-04-30 21:35:49,254 >> Configuration saved in saves/qwen3_14b/coding/fft/checkpoint-940/config.json
|
||
[INFO|configuration_utils.py:803] 2026-04-30 21:35:49,258 >> Configuration saved in saves/qwen3_14b/coding/fft/checkpoint-940/generation_config.json
|
||
|
||
Writing model shards: 0%| | 0/1 [00:00<?, ?it/s][A
|
||
Writing model shards: 100%|██████████| 1/1 [00:46<00:00, 46.35s/it][A
Writing model shards: 100%|██████████| 1/1 [00:46<00:00, 46.35s/it]
|
||
[INFO|modeling_utils.py:3380] 2026-04-30 21:36:35,646 >> Model weights saved in saves/qwen3_14b/coding/fft/checkpoint-940/model.safetensors
|
||
[INFO|tokenization_utils_base.py:3224] 2026-04-30 21:36:35,652 >> chat template saved in saves/qwen3_14b/coding/fft/checkpoint-940/chat_template.jinja
|
||
[INFO|tokenization_utils_base.py:2078] 2026-04-30 21:36:35,657 >> tokenizer config file saved in saves/qwen3_14b/coding/fft/checkpoint-940/tokenizer_config.json
|
||
[INFO|trainer.py:1863] 2026-04-30 21:36:36,733 >>
|
||
|
||
Training completed. Do not forget to share your model on huggingface.co/models =)
|
||
|
||
|
||
{'train_runtime': '1.899e+04', 'train_samples_per_second': '6.32', 'train_steps_per_second': '0.05', 'train_loss': '0.4443', 'epoch': '4'}
|
||
100%|██████████| 940/940 [5:16:26<00:00, 16.47s/it]
100%|██████████| 940/940 [5:16:26<00:00, 20.20s/it]
|
||
[INFO|trainer.py:3797] 2026-04-30 21:36:57,332 >> Saving model checkpoint to saves/qwen3_14b/coding/fft
|
||
[INFO|configuration_utils.py:432] 2026-04-30 21:36:57,341 >> Configuration saved in saves/qwen3_14b/coding/fft/config.json
|
||
[INFO|configuration_utils.py:803] 2026-04-30 21:36:57,346 >> Configuration saved in saves/qwen3_14b/coding/fft/generation_config.json
|
||
Writing model shards: 0%| | 0/1 [00:00<?, ?it/s]
Writing model shards: 100%|██████████| 1/1 [00:45<00:00, 45.33s/it]
Writing model shards: 100%|██████████| 1/1 [00:45<00:00, 45.33s/it]
|
||
[INFO|modeling_utils.py:3380] 2026-04-30 21:37:42,719 >> Model weights saved in saves/qwen3_14b/coding/fft/model.safetensors
|
||
[INFO|tokenization_utils_base.py:3224] 2026-04-30 21:37:42,726 >> chat template saved in saves/qwen3_14b/coding/fft/chat_template.jinja
|
||
[INFO|tokenization_utils_base.py:2078] 2026-04-30 21:37:42,733 >> tokenizer config file saved in saves/qwen3_14b/coding/fft/tokenizer_config.json
|
||
[INFO|trainer.py:3797] 2026-04-30 21:38:04,006 >> Saving model checkpoint to saves/qwen3_14b/coding/fft
|
||
[INFO|configuration_utils.py:432] 2026-04-30 21:38:04,013 >> Configuration saved in saves/qwen3_14b/coding/fft/config.json
|
||
[INFO|configuration_utils.py:803] 2026-04-30 21:38:04,016 >> Configuration saved in saves/qwen3_14b/coding/fft/generation_config.json
|
||
Writing model shards: 0%| | 0/1 [00:00<?, ?it/s]
Writing model shards: 100%|██████████| 1/1 [00:47<00:00, 47.06s/it]
Writing model shards: 100%|██████████| 1/1 [00:47<00:00, 47.06s/it]
|
||
[INFO|modeling_utils.py:3380] 2026-04-30 21:38:51,114 >> Model weights saved in saves/qwen3_14b/coding/fft/model.safetensors
|
||
[INFO|tokenization_utils_base.py:3224] 2026-04-30 21:38:51,121 >> chat template saved in saves/qwen3_14b/coding/fft/chat_template.jinja
|
||
[INFO|tokenization_utils_base.py:2078] 2026-04-30 21:38:51,126 >> tokenizer config file saved in saves/qwen3_14b/coding/fft/tokenizer_config.json
|
||
[INFO|modelcard.py:266] 2026-04-30 21:38:52,374 >> Dropping the following result as it does not have all the necessary fields:
|
||
{'task': {'name': 'Causal Language Modeling', 'type': 'text-generation'}}
|
||
Processing Files (0 / 0) : | | 0.00B / 0.00B
|
||
New Data Upload : | | 0.00B / 0.00B [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
Processing Files (1 / 1) : 0%| | 11.4MB / 29.6GB, 28.5MB/s
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 608kB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 608kB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 0%| | 19.3MB / 29.6GB, 16.1MB/s
|
||
New Data Upload : 0%| | 608kB / 134MB, 507kB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 1.22MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 0%| | 19.9MB / 29.6GB, 14.2MB/s
|
||
New Data Upload : 1%| | 1.22MB / 134MB, 869kB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 7.30MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 0%| | 26.0MB / 29.6GB, 16.2MB/s
|
||
New Data Upload : 5%|▌ | 7.30MB / 134MB, 4.56MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 20.1MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 0%| | 38.7MB / 29.6GB, 21.5MB/s
|
||
New Data Upload : 15%|█▍ | 20.1MB / 134MB, 11.1MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 34.0MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 0%| | 52.7MB / 29.6GB, 26.3MB/s
|
||
New Data Upload : 25%|██▌ | 34.0MB / 134MB, 17.0MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 48.0MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 0%| | 66.7MB / 29.6GB, 30.3MB/s
|
||
New Data Upload : 36%|███▌ | 48.0MB / 134MB, 21.8MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 62.0MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 0%| | 80.7MB / 29.6GB, 33.6MB/s
|
||
New Data Upload : 46%|████▌ | 62.0MB / 134MB, 25.8MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 66.9MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 0%| | 85.5MB / 29.6GB, 32.9MB/s
|
||
New Data Upload : 50%|████▉ | 66.9MB / 134MB, 25.7MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 66.9MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 66.9MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 66.9MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 67.1MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 0%| | 85.7MB / 29.6GB, 25.2MB/s
|
||
New Data Upload : 50%|█████ | 67.1MB / 134MB, 19.7MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 67.7MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 0%| | 86.3MB / 29.6GB, 24.0MB/s
|
||
New Data Upload : 34%|███▎ | 67.7MB / 201MB, 18.8MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 76.2MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 0%| | 94.8MB / 29.6GB, 25.0MB/s
|
||
New Data Upload : 28%|██▊ | 76.2MB / 268MB, 20.0MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 96.8MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 0%| | 115MB / 29.6GB, 28.9MB/s
|
||
New Data Upload : 36%|███▌ | 96.8MB / 268MB, 24.2MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 122MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 0%| | 141MB / 29.6GB, 33.5MB/s
|
||
New Data Upload : 36%|███▋ | 122MB / 335MB, 29.1MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 1%| | 154MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 1%| | 173MB / 29.6GB, 39.3MB/s
|
||
New Data Upload : 46%|████▌ | 154MB / 335MB, 35.1MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 1%| | 189MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 1%| | 208MB / 29.6GB, 45.1MB/s
|
||
New Data Upload : 47%|████▋ | 189MB / 402MB, 41.1MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 1%| | 220MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 1%| | 239MB / 29.6GB, 49.8MB/s
|
||
New Data Upload : 55%|█████▍ | 220MB / 402MB, 45.9MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 1%| | 247MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 1%| | 266MB / 29.6GB, 53.1MB/s
|
||
New Data Upload : 61%|██████▏ | 247MB / 402MB, 49.4MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 1%| | 274MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 1%| | 293MB / 29.6GB, 56.3MB/s
|
||
New Data Upload : 58%|█████▊ | 274MB / 470MB, 52.7MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 1%| | 302MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 1%| | 321MB / 29.6GB, 59.4MB/s
|
||
New Data Upload : 64%|██████▍ | 302MB / 470MB, 55.9MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 1%| | 327MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 1%| | 346MB / 29.6GB, 61.8MB/s
|
||
New Data Upload : 70%|██████▉ | 327MB / 470MB, 58.5MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 1%| | 368MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 1%|▏ | 387MB / 29.6GB, 66.6MB/s
|
||
New Data Upload : 69%|██████▊ | 368MB / 537MB, 63.4MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 1%|▏ | 401MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 1%|▏ | 419MB / 29.6GB, 69.9MB/s
|
||
New Data Upload : 66%|██████▋ | 401MB / 604MB, 66.8MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 1%|▏ | 434MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 2%|▏ | 452MB / 29.6GB, 73.0MB/s
|
||
New Data Upload : 72%|███████▏ | 434MB / 604MB, 69.9MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 2%|▏ | 486MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 2%|▏ | 504MB / 29.6GB, 78.8MB/s
|
||
New Data Upload : 72%|███████▏ | 486MB / 671MB, 75.9MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 2%|▏ | 523MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 2%|▏ | 541MB / 29.6GB, 82.0MB/s
|
||
New Data Upload : 78%|███████▊ | 523MB / 671MB, 79.2MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 2%|▏ | 551MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 2%|▏ | 570MB / 29.6GB, 83.8MB/s
|
||
New Data Upload : 75%|███████▍ | 551MB / 738MB, 81.0MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 2%|▏ | 581MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 2%|▏ | 599MB / 29.6GB, 85.6MB/s
|
||
New Data Upload : 79%|███████▊ | 581MB / 738MB, 83.0MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 2%|▏ | 631MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 2%|▏ | 649MB / 29.6GB, 90.2MB/s
|
||
New Data Upload : 78%|███████▊ | 631MB / 805MB, 87.6MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 2%|▏ | 664MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 2%|▏ | 683MB / 29.6GB, 92.2MB/s
|
||
New Data Upload : 82%|████████▏ | 664MB / 805MB, 89.7MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 2%|▏ | 705MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 2%|▏ | 724MB / 29.6GB, 95.2MB/s
|
||
New Data Upload : 81%|████████ | 705MB / 872MB, 92.8MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 3%|▎ | 746MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 3%|▎ | 764MB / 29.6GB, 98.0MB/s
|
||
New Data Upload : 79%|███████▉ | 746MB / 939MB, 95.6MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 3%|▎ | 775MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 3%|▎ | 793MB / 29.6GB, 99.2MB/s
|
||
New Data Upload : 82%|████████▏ | 775MB / 939MB, 96.8MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 3%|▎ | 802MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 3%|▎ | 820MB / 29.6GB, 100MB/s
|
||
New Data Upload : 85%|████████▌ | 802MB / 939MB, 97.8MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 3%|▎ | 831MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 3%|▎ | 850MB / 29.6GB, 101MB/s
|
||
New Data Upload : 83%|████████▎ | 831MB / 1.01GB, 99.0MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 3%|▎ | 866MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 3%|▎ | 884MB / 29.6GB, 103MB/s
|
||
New Data Upload : 86%|████████▌ | 866MB / 1.01GB, 101MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 3%|▎ | 898MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 3%|▎ | 916MB / 29.6GB, 104MB/s
|
||
New Data Upload : 84%|████████▎ | 898MB / 1.07GB, 102MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 3%|▎ | 941MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 3%|▎ | 960MB / 29.6GB, 107MB/s
|
||
New Data Upload : 83%|████████▎ | 941MB / 1.14GB, 105MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 3%|▎ | 986MB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 3%|▎ | 1.01GB / 29.6GB, 109MB/s
|
||
New Data Upload : 87%|████████▋ | 986MB / 1.14GB, 107MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 3%|▎ | 1.01GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 3%|▎ | 1.03GB / 29.6GB, 110MB/s
|
||
New Data Upload : 84%|████████▍ | 1.01GB / 1.21GB, 108MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 4%|▎ | 1.04GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 4%|▎ | 1.06GB / 29.6GB, 111MB/s
|
||
New Data Upload : 86%|████████▋ | 1.04GB / 1.21GB, 109MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 4%|▎ | 1.07GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 4%|▎ | 1.09GB / 29.6GB, 111MB/s
|
||
New Data Upload : 84%|████████▍ | 1.07GB / 1.27GB, 110MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 4%|▍ | 1.11GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 4%|▍ | 1.13GB / 29.6GB, 113MB/s
|
||
New Data Upload : 87%|████████▋ | 1.11GB / 1.27GB, 111MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 4%|▍ | 1.14GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 4%|▍ | 1.16GB / 29.6GB, 114MB/s
|
||
New Data Upload : 85%|████████▌ | 1.14GB / 1.34GB, 112MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 4%|▍ | 1.18GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 4%|▍ | 1.20GB / 29.6GB, 117MB/s
|
||
New Data Upload : 84%|████████▎ | 1.18GB / 1.41GB, 115MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 4%|▍ | 1.23GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 4%|▍ | 1.24GB / 29.6GB, 121MB/s
|
||
New Data Upload : 87%|████████▋ | 1.23GB / 1.41GB, 120MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 4%|▍ | 1.28GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 4%|▍ | 1.30GB / 29.6GB, 127MB/s
|
||
New Data Upload : 87%|████████▋ | 1.28GB / 1.48GB, 126MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 4%|▍ | 1.33GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 5%|▍ | 1.34GB / 29.6GB, 131MB/s
|
||
New Data Upload : 86%|████████▌ | 1.33GB / 1.54GB, 130MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 5%|▍ | 1.36GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 5%|▍ | 1.38GB / 29.6GB, 134MB/s
|
||
New Data Upload : 88%|████████▊ | 1.36GB / 1.54GB, 133MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 5%|▍ | 1.39GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 5%|▍ | 1.41GB / 29.6GB, 136MB/s
|
||
New Data Upload : 86%|████████▌ | 1.39GB / 1.61GB, 136MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 5%|▍ | 1.42GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 5%|▍ | 1.44GB / 29.6GB, 139MB/s
|
||
New Data Upload : 88%|████████▊ | 1.42GB / 1.61GB, 139MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 5%|▍ | 1.46GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 5%|▌ | 1.48GB / 29.6GB, 143MB/s
|
||
New Data Upload : 87%|████████▋ | 1.46GB / 1.68GB, 143MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 5%|▌ | 1.51GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 5%|▌ | 1.53GB / 29.6GB, 146MB/s
|
||
New Data Upload : 90%|█████████ | 1.51GB / 1.68GB, 146MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 5%|▌ | 1.55GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 5%|▌ | 1.57GB / 29.6GB, 149MB/s
|
||
New Data Upload : 89%|████████▉ | 1.55GB / 1.74GB, 149MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 5%|▌ | 1.59GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 5%|▌ | 1.61GB / 29.6GB, 151MB/s
|
||
New Data Upload : 88%|████████▊ | 1.59GB / 1.81GB, 151MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 6%|▌ | 1.63GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 6%|▌ | 1.65GB / 29.6GB, 154MB/s
|
||
New Data Upload : 90%|█████████ | 1.63GB / 1.81GB, 154MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 6%|▌ | 1.68GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 6%|▌ | 1.70GB / 29.6GB, 158MB/s
|
||
New Data Upload : 89%|████████▉ | 1.68GB / 1.88GB, 158MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 6%|▌ | 1.72GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 6%|▌ | 1.74GB / 29.6GB, 162MB/s
|
||
New Data Upload : 91%|█████████▏| 1.72GB / 1.88GB, 162MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 6%|▌ | 1.76GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 6%|▌ | 1.78GB / 29.6GB, 166MB/s
|
||
New Data Upload : 91%|█████████ | 1.76GB / 1.94GB, 166MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 6%|▌ | 1.80GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 6%|▌ | 1.82GB / 29.6GB, 170MB/s
|
||
New Data Upload : 93%|█████████▎| 1.80GB / 1.94GB, 170MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 6%|▌ | 1.83GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 6%|▋ | 1.85GB / 29.6GB, 173MB/s
|
||
New Data Upload : 91%|█████████ | 1.83GB / 2.01GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 6%|▋ | 1.88GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 6%|▋ | 1.89GB / 29.6GB, 177MB/s
|
||
New Data Upload : 93%|█████████▎| 1.87GB / 2.01GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 6%|▋ | 1.90GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 6%|▋ | 1.92GB / 29.6GB, 179MB/s
|
||
New Data Upload : 91%|█████████▏| 1.90GB / 2.08GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 7%|▋ | 1.92GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 7%|▋ | 1.94GB / 29.6GB, 179MB/s
|
||
New Data Upload : 92%|█████████▏| 1.92GB / 2.08GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 7%|▋ | 1.95GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 7%|▋ | 1.97GB / 29.6GB, 180MB/s
|
||
New Data Upload : 94%|█████████▍| 1.95GB / 2.08GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 7%|▋ | 1.99GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 7%|▋ | 2.01GB / 29.6GB, 180MB/s
|
||
New Data Upload : 93%|█████████▎| 1.99GB / 2.15GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 7%|▋ | 2.03GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 7%|▋ | 2.04GB / 29.6GB, 180MB/s
|
||
New Data Upload : 91%|█████████▏| 2.02GB / 2.21GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 7%|▋ | 2.06GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 7%|▋ | 2.08GB / 29.6GB, 180MB/s
|
||
New Data Upload : 93%|█████████▎| 2.06GB / 2.21GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 7%|▋ | 2.11GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 7%|▋ | 2.13GB / 29.6GB, 182MB/s
|
||
New Data Upload : 92%|█████████▏| 2.10GB / 2.28GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 7%|▋ | 2.13GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 7%|▋ | 2.15GB / 29.6GB, 182MB/s
|
||
New Data Upload : 93%|█████████▎| 2.12GB / 2.28GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 7%|▋ | 2.14GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 7%|▋ | 2.16GB / 29.6GB, 181MB/s
|
||
New Data Upload : 94%|█████████▍| 2.14GB / 2.28GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 7%|▋ | 2.17GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 7%|▋ | 2.19GB / 29.6GB, 181MB/s
|
||
New Data Upload : 92%|█████████▏| 2.17GB / 2.35GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 7%|▋ | 2.20GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 7%|▋ | 2.22GB / 29.6GB, 179MB/s
|
||
New Data Upload : 93%|█████████▎| 2.19GB / 2.35GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 8%|▊ | 2.23GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 8%|▊ | 2.25GB / 29.6GB, 180MB/s
|
||
New Data Upload : 92%|█████████▏| 2.23GB / 2.41GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 8%|▊ | 2.28GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 8%|▊ | 2.30GB / 29.6GB, 181MB/s
|
||
New Data Upload : 92%|█████████▏| 2.27GB / 2.48GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 8%|▊ | 2.31GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 8%|▊ | 2.33GB / 29.6GB, 179MB/s
|
||
New Data Upload : 93%|█████████▎| 2.30GB / 2.48GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 8%|▊ | 2.35GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 8%|▊ | 2.37GB / 29.6GB, 179MB/s
|
||
New Data Upload : 92%|█████████▏| 2.34GB / 2.55GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 8%|▊ | 2.42GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 8%|▊ | 2.44GB / 29.6GB, 183MB/s
|
||
New Data Upload : 94%|█████████▍| 2.39GB / 2.55GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 8%|▊ | 2.50GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 9%|▊ | 2.52GB / 29.6GB, 188MB/s
|
||
New Data Upload : 93%|█████████▎| 2.43GB / 2.62GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 9%|▉ | 2.62GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 9%|▉ | 2.64GB / 29.6GB, 195MB/s
|
||
New Data Upload : 95%|█████████▍| 2.47GB / 2.62GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 9%|▉ | 2.76GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 9%|▉ | 2.78GB / 29.6GB, 205MB/s
|
||
New Data Upload : 96%|█████████▌| 2.52GB / 2.62GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 10%|▉ | 2.93GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 10%|▉ | 2.94GB / 29.6GB, 218MB/s
|
||
New Data Upload : 98%|█████████▊| 2.55GB / 2.62GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 10%|█ | 2.96GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 10%|█ | 2.98GB / 29.6GB, 217MB/s
|
||
New Data Upload : 96%|█████████▋| 2.58GB / 2.68GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 10%|█ | 3.02GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 10%|█ | 3.04GB / 29.6GB, 220MB/s
|
||
New Data Upload : 97%|█████████▋| 2.60GB / 2.68GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 10%|█ | 3.03GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 10%|█ | 3.05GB / 29.6GB, 219MB/s
|
||
New Data Upload : 95%|█████████▌| 2.62GB / 2.75GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 10%|█ | 3.04GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 10%|█ | 3.06GB / 29.6GB, 216MB/s
|
||
New Data Upload : 95%|█████████▌| 2.62GB / 2.75GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 10%|█ | 3.06GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 10%|█ | 3.08GB / 29.6GB, 215MB/s
|
||
New Data Upload : 94%|█████████▍| 2.64GB / 2.82GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 10%|█ | 3.08GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 10%|█ | 3.10GB / 29.6GB, 214MB/s
|
||
New Data Upload : 95%|█████████▍| 2.67GB / 2.82GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 11%|█ | 3.11GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 11%|█ | 3.13GB / 29.6GB, 213MB/s
|
||
New Data Upload : 94%|█████████▎| 2.70GB / 2.88GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 11%|█ | 3.14GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 11%|█ | 3.16GB / 29.6GB, 212MB/s
|
||
New Data Upload : 95%|█████████▍| 2.73GB / 2.88GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 11%|█ | 3.18GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 11%|█ | 3.20GB / 29.6GB, 212MB/s
|
||
New Data Upload : 94%|█████████▎| 2.76GB / 2.95GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 11%|█ | 3.21GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 11%|█ | 3.23GB / 29.6GB, 212MB/s
|
||
New Data Upload : 93%|█████████▎| 2.79GB / 3.02GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 11%|█ | 3.24GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 11%|█ | 3.26GB / 29.6GB, 212MB/s
|
||
New Data Upload : 94%|█████████▎| 2.83GB / 3.02GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 11%|█ | 3.28GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 11%|█ | 3.30GB / 29.6GB, 213MB/s
|
||
New Data Upload : 93%|█████████▎| 2.87GB / 3.08GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 11%|█▏ | 3.33GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 11%|█▏ | 3.34GB / 29.6GB, 214MB/s
|
||
New Data Upload : 94%|█████████▍| 2.91GB / 3.08GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 11%|█▏ | 3.36GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 11%|█▏ | 3.38GB / 29.6GB, 214MB/s
|
||
New Data Upload : 94%|█████████▎| 2.95GB / 3.15GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 12%|█▏ | 3.40GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 12%|█▏ | 3.42GB / 29.6GB, 213MB/s
|
||
New Data Upload : 95%|█████████▍| 2.98GB / 3.15GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 12%|█▏ | 3.43GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 12%|█▏ | 3.45GB / 29.6GB, 210MB/s
|
||
New Data Upload : 94%|█████████▎| 3.02GB / 3.22GB, 170MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 12%|█▏ | 3.48GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 12%|█▏ | 3.50GB / 29.6GB, 211MB/s
|
||
New Data Upload : 95%|█████████▌| 3.06GB / 3.22GB, 170MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 12%|█▏ | 3.52GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 12%|█▏ | 3.54GB / 29.6GB, 212MB/s
|
||
New Data Upload : 94%|█████████▍| 3.10GB / 3.29GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 12%|█▏ | 3.55GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 12%|█▏ | 3.57GB / 29.6GB, 212MB/s
|
||
New Data Upload : 93%|█████████▎| 3.13GB / 3.35GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 12%|█▏ | 3.58GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 12%|█▏ | 3.60GB / 29.6GB, 212MB/s
|
||
New Data Upload : 94%|█████████▍| 3.17GB / 3.35GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 12%|█▏ | 3.62GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 12%|█▏ | 3.64GB / 29.6GB, 211MB/s
|
||
New Data Upload : 94%|█████████▎| 3.20GB / 3.42GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 12%|█▏ | 3.64GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 12%|█▏ | 3.66GB / 29.6GB, 209MB/s
|
||
New Data Upload : 94%|█████████▍| 3.22GB / 3.42GB, 168MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 12%|█▏ | 3.67GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 12%|█▏ | 3.69GB / 29.6GB, 208MB/s
|
||
New Data Upload : 93%|█████████▎| 3.26GB / 3.49GB, 167MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 13%|█▎ | 3.73GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 13%|█▎ | 3.75GB / 29.6GB, 210MB/s
|
||
New Data Upload : 95%|█████████▍| 3.31GB / 3.49GB, 169MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 13%|█▎ | 3.77GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 13%|█▎ | 3.79GB / 29.6GB, 209MB/s
|
||
New Data Upload : 94%|█████████▍| 3.35GB / 3.55GB, 168MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 13%|█▎ | 3.82GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 13%|█▎ | 3.84GB / 29.6GB, 210MB/s
|
||
New Data Upload : 94%|█████████▍| 3.40GB / 3.62GB, 169MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 13%|█▎ | 3.87GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 13%|█▎ | 3.89GB / 29.6GB, 211MB/s
|
||
New Data Upload : 95%|█████████▌| 3.45GB / 3.62GB, 170MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 13%|█▎ | 3.90GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 13%|█▎ | 3.92GB / 29.6GB, 210MB/s
|
||
New Data Upload : 95%|█████████▍| 3.49GB / 3.69GB, 169MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 13%|█▎ | 3.93GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 13%|█▎ | 3.95GB / 29.6GB, 209MB/s
|
||
New Data Upload : 95%|█████████▌| 3.52GB / 3.69GB, 168MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 13%|█▎ | 3.97GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 13%|█▎ | 3.99GB / 29.6GB, 210MB/s
|
||
New Data Upload : 95%|█████████▍| 3.56GB / 3.76GB, 169MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 14%|█▎ | 4.00GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 14%|█▎ | 4.02GB / 29.6GB, 208MB/s
|
||
New Data Upload : 95%|█████████▌| 3.58GB / 3.76GB, 168MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 14%|█▎ | 4.03GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 14%|█▎ | 4.05GB / 29.6GB, 208MB/s
|
||
New Data Upload : 94%|█████████▍| 3.61GB / 3.82GB, 168MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 14%|█▎ | 4.06GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 14%|█▍ | 4.08GB / 29.6GB, 209MB/s
|
||
New Data Upload : 95%|█████████▌| 3.64GB / 3.82GB, 169MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 14%|█▍ | 4.11GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 14%|█▍ | 4.12GB / 29.6GB, 211MB/s
|
||
New Data Upload : 95%|█████████▍| 3.69GB / 3.89GB, 170MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 14%|█▍ | 4.15GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 14%|█▍ | 4.17GB / 29.6GB, 212MB/s
|
||
New Data Upload : 94%|█████████▍| 3.73GB / 3.96GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 14%|█▍ | 4.18GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 14%|█▍ | 4.20GB / 29.6GB, 211MB/s
|
||
New Data Upload : 95%|█████████▌| 3.77GB / 3.96GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 14%|█▍ | 4.22GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 14%|█▍ | 4.24GB / 29.6GB, 212MB/s
|
||
New Data Upload : 95%|█████████▍| 3.81GB / 4.02GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 14%|█▍ | 4.26GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 14%|█▍ | 4.28GB / 29.6GB, 211MB/s
|
||
New Data Upload : 96%|█████████▌| 3.85GB / 4.02GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 15%|█▍ | 4.29GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 15%|█▍ | 4.31GB / 29.6GB, 213MB/s
|
||
New Data Upload : 95%|█████████▍| 3.88GB / 4.09GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 15%|█▍ | 4.33GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 15%|█▍ | 4.35GB / 29.6GB, 215MB/s
|
||
New Data Upload : 94%|█████████▍| 3.92GB / 4.16GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 15%|█▍ | 4.37GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 15%|█▍ | 4.39GB / 29.6GB, 216MB/s
|
||
New Data Upload : 95%|█████████▌| 3.96GB / 4.16GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 15%|█▍ | 4.42GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 15%|█▌ | 4.44GB / 29.6GB, 218MB/s
|
||
New Data Upload : 96%|█████████▋| 4.01GB / 4.16GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 15%|█▌ | 4.46GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 15%|█▌ | 4.48GB / 29.6GB, 219MB/s
|
||
New Data Upload : 96%|█████████▌| 4.05GB / 4.22GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 15%|█▌ | 4.50GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 15%|█▌ | 4.52GB / 29.6GB, 218MB/s
|
||
New Data Upload : 95%|█████████▌| 4.09GB / 4.29GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 15%|█▌ | 4.53GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 15%|█▌ | 4.55GB / 29.6GB, 218MB/s
|
||
New Data Upload : 96%|█████████▌| 4.11GB / 4.29GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 15%|█▌ | 4.57GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 16%|█▌ | 4.59GB / 29.6GB, 217MB/s
|
||
New Data Upload : 95%|█████████▌| 4.15GB / 4.36GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 16%|█▌ | 4.60GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 16%|█▌ | 4.62GB / 29.6GB, 214MB/s
|
||
New Data Upload : 96%|█████████▌| 4.18GB / 4.36GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 16%|█▌ | 4.64GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 16%|█▌ | 4.65GB / 29.6GB, 209MB/s
|
||
New Data Upload : 95%|█████████▌| 4.22GB / 4.43GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 16%|█▌ | 4.67GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 16%|█▌ | 4.69GB / 29.6GB, 201MB/s
|
||
New Data Upload : 96%|█████████▌| 4.26GB / 4.43GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 16%|█▌ | 4.71GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 16%|█▌ | 4.73GB / 29.6GB, 191MB/s
|
||
New Data Upload : 96%|█████████▌| 4.29GB / 4.49GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 16%|█▌ | 4.76GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 16%|█▌ | 4.78GB / 29.6GB, 180MB/s
|
||
New Data Upload : 95%|█████████▌| 4.34GB / 4.56GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 16%|█▋ | 4.81GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 16%|█▋ | 4.83GB / 29.6GB, 182MB/s
|
||
New Data Upload : 96%|█████████▋| 4.40GB / 4.56GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 16%|█▋ | 4.85GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 16%|█▋ | 4.87GB / 29.6GB, 180MB/s
|
||
New Data Upload : 96%|█████████▌| 4.43GB / 4.63GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 17%|█▋ | 4.88GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 17%|█▋ | 4.90GB / 29.6GB, 181MB/s
|
||
New Data Upload : 95%|█████████▌| 4.46GB / 4.69GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 17%|█▋ | 4.92GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 17%|█▋ | 4.93GB / 29.6GB, 184MB/s
|
||
New Data Upload : 96%|█████████▌| 4.50GB / 4.69GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 17%|█▋ | 4.96GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 17%|█▋ | 4.97GB / 29.6GB, 186MB/s
|
||
New Data Upload : 95%|█████████▌| 4.54GB / 4.76GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 17%|█▋ | 5.00GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 17%|█▋ | 5.01GB / 29.6GB, 187MB/s
|
||
New Data Upload : 95%|█████████▍| 4.58GB / 4.83GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 17%|█▋ | 5.04GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 17%|█▋ | 5.06GB / 29.6GB, 189MB/s
|
||
New Data Upload : 96%|█████████▌| 4.62GB / 4.83GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 17%|█▋ | 5.10GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 17%|█▋ | 5.11GB / 29.6GB, 191MB/s
|
||
New Data Upload : 96%|█████████▌| 4.68GB / 4.90GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 17%|█▋ | 5.16GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 18%|█▊ | 5.18GB / 29.6GB, 194MB/s
|
||
New Data Upload : 97%|█████████▋| 4.74GB / 4.90GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 18%|█▊ | 5.19GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 18%|█▊ | 5.21GB / 29.6GB, 194MB/s
|
||
New Data Upload : 96%|█████████▌| 4.78GB / 4.96GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 18%|█▊ | 5.22GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 18%|█▊ | 5.24GB / 29.6GB, 194MB/s
|
||
New Data Upload : 97%|█████████▋| 4.81GB / 4.96GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 18%|█▊ | 5.25GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 18%|█▊ | 5.27GB / 29.6GB, 193MB/s
|
||
New Data Upload : 96%|█████████▌| 4.83GB / 5.03GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 18%|█▊ | 5.29GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 18%|█▊ | 5.31GB / 29.6GB, 192MB/s
|
||
New Data Upload : 97%|█████████▋| 4.87GB / 5.03GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 18%|█▊ | 5.32GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 18%|█▊ | 5.34GB / 29.6GB, 192MB/s
|
||
New Data Upload : 96%|█████████▌| 4.90GB / 5.10GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 18%|█▊ | 5.35GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 18%|█▊ | 5.37GB / 29.6GB, 192MB/s
|
||
New Data Upload : 96%|█████████▌| 4.94GB / 5.16GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 18%|█▊ | 5.40GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 18%|█▊ | 5.42GB / 29.6GB, 193MB/s
|
||
New Data Upload : 96%|█████████▋| 4.98GB / 5.16GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 18%|█▊ | 5.45GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 18%|█▊ | 5.47GB / 29.6GB, 193MB/s
|
||
New Data Upload : 96%|█████████▌| 5.03GB / 5.23GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 19%|█▊ | 5.49GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 19%|█▊ | 5.51GB / 29.6GB, 193MB/s
|
||
New Data Upload : 97%|█████████▋| 5.07GB / 5.23GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 19%|█▊ | 5.53GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 19%|█▉ | 5.55GB / 29.6GB, 194MB/s
|
||
New Data Upload : 97%|█████████▋| 5.12GB / 5.30GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 19%|█▉ | 5.56GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 19%|█▉ | 5.57GB / 29.6GB, 193MB/s
|
||
New Data Upload : 97%|█████████▋| 5.14GB / 5.30GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 19%|█▉ | 5.58GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 19%|█▉ | 5.60GB / 29.6GB, 192MB/s
|
||
New Data Upload : 96%|█████████▋| 5.16GB / 5.36GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 19%|█▉ | 5.61GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 19%|█▉ | 5.63GB / 29.6GB, 193MB/s
|
||
New Data Upload : 96%|█████████▌| 5.19GB / 5.43GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 19%|█▉ | 5.64GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 19%|█▉ | 5.66GB / 29.6GB, 193MB/s
|
||
New Data Upload : 96%|█████████▌| 5.23GB / 5.43GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 19%|█▉ | 5.67GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 19%|█▉ | 5.69GB / 29.6GB, 191MB/s
|
||
New Data Upload : 96%|█████████▌| 5.26GB / 5.50GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 19%|█▉ | 5.71GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 19%|█▉ | 5.73GB / 29.6GB, 191MB/s
|
||
New Data Upload : 96%|█████████▋| 5.30GB / 5.50GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 20%|█▉ | 5.77GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 20%|█▉ | 5.78GB / 29.6GB, 191MB/s
|
||
New Data Upload : 96%|█████████▌| 5.35GB / 5.57GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 20%|█▉ | 5.81GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 20%|█▉ | 5.83GB / 29.6GB, 190MB/s
|
||
New Data Upload : 97%|█████████▋| 5.39GB / 5.57GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 20%|█▉ | 5.85GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 20%|█▉ | 5.87GB / 29.6GB, 191MB/s
|
||
New Data Upload : 97%|█████████▋| 5.44GB / 5.63GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 20%|█▉ | 5.90GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 20%|██ | 5.92GB / 29.6GB, 193MB/s
|
||
New Data Upload : 97%|█████████▋| 5.48GB / 5.63GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 20%|██ | 5.94GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 20%|██ | 5.96GB / 29.6GB, 193MB/s
|
||
New Data Upload : 97%|█████████▋| 5.52GB / 5.70GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 20%|██ | 5.97GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 20%|██ | 5.99GB / 29.6GB, 193MB/s
|
||
New Data Upload : 96%|█████████▋| 5.55GB / 5.77GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 20%|██ | 6.00GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 20%|██ | 6.01GB / 29.6GB, 193MB/s
|
||
New Data Upload : 97%|█████████▋| 5.58GB / 5.77GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 20%|██ | 6.03GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 20%|██ | 6.05GB / 29.6GB, 193MB/s
|
||
New Data Upload : 96%|█████████▌| 5.61GB / 5.83GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 21%|██ | 6.06GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 21%|██ | 6.08GB / 29.6GB, 192MB/s
|
||
New Data Upload : 96%|█████████▌| 5.65GB / 5.90GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 21%|██ | 6.10GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 21%|██ | 6.12GB / 29.6GB, 191MB/s
|
||
New Data Upload : 96%|█████████▋| 5.68GB / 5.90GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 21%|██ | 6.14GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 21%|██ | 6.16GB / 29.6GB, 192MB/s
|
||
New Data Upload : 97%|█████████▋| 5.72GB / 5.90GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 21%|██ | 6.18GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 21%|██ | 6.20GB / 29.6GB, 192MB/s
|
||
New Data Upload : 97%|█████████▋| 5.77GB / 5.97GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 21%|██ | 6.22GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 21%|██ | 6.24GB / 29.6GB, 192MB/s
|
||
New Data Upload : 96%|█████████▌| 5.81GB / 6.04GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 21%|██ | 6.26GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 21%|██ | 6.28GB / 29.6GB, 193MB/s
|
||
New Data Upload : 97%|█████████▋| 5.85GB / 6.04GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 21%|██▏ | 6.30GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 21%|██▏ | 6.32GB / 29.6GB, 193MB/s
|
||
New Data Upload : 96%|█████████▋| 5.89GB / 6.10GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 21%|██▏ | 6.34GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 21%|██▏ | 6.36GB / 29.6GB, 193MB/s
|
||
New Data Upload : 97%|█████████▋| 5.92GB / 6.10GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 22%|██▏ | 6.37GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 22%|██▏ | 6.39GB / 29.6GB, 191MB/s
|
||
New Data Upload : 97%|█████████▋| 5.96GB / 6.17GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 22%|██▏ | 6.42GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 22%|██▏ | 6.44GB / 29.6GB, 192MB/s
|
||
New Data Upload : 97%|█████████▋| 6.01GB / 6.17GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 22%|██▏ | 6.46GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 22%|██▏ | 6.48GB / 29.6GB, 192MB/s
|
||
New Data Upload : 97%|█████████▋| 6.04GB / 6.24GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 22%|██▏ | 6.50GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 22%|██▏ | 6.52GB / 29.6GB, 193MB/s
|
||
New Data Upload : 97%|█████████▋| 6.08GB / 6.30GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 22%|██▏ | 6.54GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 22%|██▏ | 6.56GB / 29.6GB, 194MB/s
|
||
New Data Upload : 97%|█████████▋| 6.13GB / 6.30GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 22%|██▏ | 6.58GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 22%|██▏ | 6.60GB / 29.6GB, 194MB/s
|
||
New Data Upload : 97%|█████████▋| 6.17GB / 6.37GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 22%|██▏ | 6.61GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 22%|██▏ | 6.62GB / 29.6GB, 193MB/s
|
||
New Data Upload : 97%|█████████▋| 6.19GB / 6.37GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 23%|██▎ | 6.65GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 23%|██▎ | 6.66GB / 29.6GB, 193MB/s
|
||
New Data Upload : 97%|█████████▋| 6.23GB / 6.44GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 23%|██▎ | 6.69GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 23%|██▎ | 6.71GB / 29.6GB, 195MB/s
|
||
New Data Upload : 98%|█████████▊| 6.28GB / 6.44GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 23%|██▎ | 6.72GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 23%|██▎ | 6.74GB / 29.6GB, 192MB/s
|
||
New Data Upload : 97%|█████████▋| 6.30GB / 6.50GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 23%|██▎ | 6.76GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 23%|██▎ | 6.78GB / 29.6GB, 191MB/s
|
||
New Data Upload : 96%|█████████▋| 6.34GB / 6.57GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 23%|██▎ | 6.79GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 23%|██▎ | 6.81GB / 29.6GB, 191MB/s
|
||
New Data Upload : 97%|█████████▋| 6.38GB / 6.57GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 23%|██▎ | 6.83GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 23%|██▎ | 6.85GB / 29.6GB, 191MB/s
|
||
New Data Upload : 97%|█████████▋| 6.41GB / 6.64GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 23%|██▎ | 6.87GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 23%|██▎ | 6.89GB / 29.6GB, 191MB/s
|
||
New Data Upload : 97%|█████████▋| 6.45GB / 6.64GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 23%|██▎ | 6.92GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 23%|██▎ | 6.93GB / 29.6GB, 192MB/s
|
||
New Data Upload : 97%|█████████▋| 6.50GB / 6.71GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 24%|██▎ | 6.97GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 24%|██▎ | 6.99GB / 29.6GB, 194MB/s
|
||
New Data Upload : 97%|█████████▋| 6.55GB / 6.77GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 24%|██▎ | 7.01GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 24%|██▍ | 7.03GB / 29.6GB, 193MB/s
|
||
New Data Upload : 97%|█████████▋| 6.60GB / 6.77GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 24%|██▍ | 7.06GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 24%|██▍ | 7.08GB / 29.6GB, 192MB/s
|
||
New Data Upload : 97%|█████████▋| 6.64GB / 6.84GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 24%|██▍ | 7.09GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 24%|██▍ | 7.11GB / 29.6GB, 189MB/s
|
||
New Data Upload : 97%|█████████▋| 6.67GB / 6.91GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 24%|██▍ | 7.12GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 24%|██▍ | 7.14GB / 29.6GB, 189MB/s
|
||
New Data Upload : 97%|█████████▋| 6.70GB / 6.91GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 24%|██▍ | 7.15GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 24%|██▍ | 7.17GB / 29.6GB, 189MB/s
|
||
New Data Upload : 97%|█████████▋| 6.74GB / 6.97GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 24%|██▍ | 7.19GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 24%|██▍ | 7.21GB / 29.6GB, 190MB/s
|
||
New Data Upload : 97%|█████████▋| 6.77GB / 6.97GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 24%|██▍ | 7.23GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 25%|██▍ | 7.25GB / 29.6GB, 190MB/s
|
||
New Data Upload : 97%|█████████▋| 6.81GB / 7.04GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 25%|██▍ | 7.28GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 25%|██▍ | 7.30GB / 29.6GB, 193MB/s
|
||
New Data Upload : 98%|█████████▊| 6.87GB / 7.04GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 25%|██▍ | 7.33GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 25%|██▍ | 7.35GB / 29.6GB, 194MB/s
|
||
New Data Upload : 97%|█████████▋| 6.91GB / 7.11GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 25%|██▍ | 7.38GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 25%|██▌ | 7.40GB / 29.6GB, 194MB/s
|
||
New Data Upload : 97%|█████████▋| 6.96GB / 7.18GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 25%|██▌ | 7.42GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 25%|██▌ | 7.44GB / 29.6GB, 193MB/s
|
||
New Data Upload : 98%|█████████▊| 7.00GB / 7.18GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 25%|██▌ | 7.45GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 25%|██▌ | 7.47GB / 29.6GB, 193MB/s
|
||
New Data Upload : 98%|█████████▊| 7.04GB / 7.18GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 25%|██▌ | 7.47GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 25%|██▌ | 7.49GB / 29.6GB, 190MB/s
|
||
New Data Upload : 97%|█████████▋| 7.06GB / 7.24GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 25%|██▌ | 7.49GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 25%|██▌ | 7.51GB / 29.6GB, 190MB/s
|
||
New Data Upload : 97%|█████████▋| 7.08GB / 7.31GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 25%|██▌ | 7.51GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 25%|██▌ | 7.53GB / 29.6GB, 189MB/s
|
||
New Data Upload : 97%|█████████▋| 7.10GB / 7.31GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 26%|██▌ | 7.55GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 26%|██▌ | 7.57GB / 29.6GB, 190MB/s
|
||
New Data Upload : 97%|█████████▋| 7.13GB / 7.38GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 26%|██▌ | 7.58GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 26%|██▌ | 7.60GB / 29.6GB, 190MB/s
|
||
New Data Upload : 97%|█████████▋| 7.16GB / 7.38GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 26%|██▌ | 7.62GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 26%|██▌ | 7.64GB / 29.6GB, 191MB/s
|
||
New Data Upload : 97%|█████████▋| 7.20GB / 7.44GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 26%|██▌ | 7.66GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 26%|██▌ | 7.68GB / 29.6GB, 191MB/s
|
||
New Data Upload : 96%|█████████▋| 7.24GB / 7.51GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 26%|██▌ | 7.70GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 26%|██▌ | 7.71GB / 29.6GB, 189MB/s
|
||
New Data Upload : 97%|█████████▋| 7.28GB / 7.51GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 26%|██▌ | 7.75GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 26%|██▋ | 7.76GB / 29.6GB, 190MB/s
|
||
New Data Upload : 97%|█████████▋| 7.33GB / 7.58GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 26%|██▋ | 7.79GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 26%|██▋ | 7.81GB / 29.6GB, 190MB/s
|
||
New Data Upload : 97%|█████████▋| 7.37GB / 7.58GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 27%|██▋ | 7.84GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 27%|██▋ | 7.86GB / 29.6GB, 190MB/s
|
||
New Data Upload : 97%|█████████▋| 7.43GB / 7.64GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 27%|██▋ | 7.89GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 27%|██▋ | 7.91GB / 29.6GB, 191MB/s
|
||
New Data Upload : 97%|█████████▋| 7.47GB / 7.71GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 27%|██▋ | 7.92GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 27%|██▋ | 7.94GB / 29.6GB, 191MB/s
|
||
New Data Upload : 97%|█████████▋| 7.51GB / 7.71GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 27%|██▋ | 7.97GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 27%|██▋ | 7.99GB / 29.6GB, 193MB/s
|
||
New Data Upload : 98%|█████████▊| 7.55GB / 7.71GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 27%|██▋ | 8.00GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 27%|██▋ | 8.01GB / 29.6GB, 193MB/s
|
||
New Data Upload : 97%|█████████▋| 7.58GB / 7.78GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 27%|██▋ | 8.03GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 27%|██▋ | 8.05GB / 29.6GB, 193MB/s
|
||
New Data Upload : 98%|█████████▊| 7.61GB / 7.78GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 27%|██▋ | 8.06GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 27%|██▋ | 8.08GB / 29.6GB, 193MB/s
|
||
New Data Upload : 97%|█████████▋| 7.65GB / 7.85GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 27%|██▋ | 8.10GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 27%|██▋ | 8.12GB / 29.6GB, 192MB/s
|
||
New Data Upload : 97%|█████████▋| 7.69GB / 7.91GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 28%|██▊ | 8.14GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 28%|██▊ | 8.16GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 7.72GB / 7.91GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 28%|██▊ | 8.18GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 28%|██▊ | 8.19GB / 29.6GB, 192MB/s
|
||
New Data Upload : 97%|█████████▋| 7.76GB / 7.98GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 28%|██▊ | 8.21GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 28%|██▊ | 8.23GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 7.79GB / 7.98GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 28%|██▊ | 8.25GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 28%|██▊ | 8.27GB / 29.6GB, 191MB/s
|
||
New Data Upload : 97%|█████████▋| 7.83GB / 8.05GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 28%|██▊ | 8.30GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 28%|██▊ | 8.32GB / 29.6GB, 192MB/s
|
||
New Data Upload : 97%|█████████▋| 7.88GB / 8.11GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 28%|██▊ | 8.34GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 28%|██▊ | 8.36GB / 29.6GB, 193MB/s
|
||
New Data Upload : 98%|█████████▊| 7.93GB / 8.11GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 28%|██▊ | 8.40GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 28%|██▊ | 8.41GB / 29.6GB, 194MB/s
|
||
New Data Upload : 98%|█████████▊| 7.98GB / 8.18GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 29%|██▊ | 8.45GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 29%|██▊ | 8.47GB / 29.6GB, 195MB/s
|
||
New Data Upload : 98%|█████████▊| 8.03GB / 8.18GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 29%|██▊ | 8.48GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 29%|██▊ | 8.50GB / 29.6GB, 194MB/s
|
||
New Data Upload : 98%|█████████▊| 8.06GB / 8.25GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 29%|██▉ | 8.51GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 29%|██▉ | 8.53GB / 29.6GB, 192MB/s
|
||
New Data Upload : 97%|█████████▋| 8.09GB / 8.32GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 29%|██▉ | 8.55GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 29%|██▉ | 8.57GB / 29.6GB, 193MB/s
|
||
New Data Upload : 98%|█████████▊| 8.13GB / 8.32GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 29%|██▉ | 8.59GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 29%|██▉ | 8.61GB / 29.6GB, 194MB/s
|
||
New Data Upload : 97%|█████████▋| 8.17GB / 8.38GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 29%|██▉ | 8.63GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 29%|██▉ | 8.65GB / 29.6GB, 194MB/s
|
||
New Data Upload : 98%|█████████▊| 8.21GB / 8.38GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 29%|██▉ | 8.66GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 29%|██▉ | 8.68GB / 29.6GB, 193MB/s
|
||
New Data Upload : 98%|█████████▊| 8.25GB / 8.45GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 29%|██▉ | 8.70GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 29%|██▉ | 8.72GB / 29.6GB, 194MB/s
|
||
New Data Upload : 98%|█████████▊| 8.28GB / 8.45GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 30%|██▉ | 8.74GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 30%|██▉ | 8.76GB / 29.6GB, 194MB/s
|
||
New Data Upload : 98%|█████████▊| 8.33GB / 8.52GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 30%|██▉ | 8.79GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 30%|██▉ | 8.81GB / 29.6GB, 196MB/s
|
||
New Data Upload : 98%|█████████▊| 8.37GB / 8.58GB, 196MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 30%|██▉ | 8.82GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 30%|██▉ | 8.84GB / 29.6GB, 195MB/s
|
||
New Data Upload : 98%|█████████▊| 8.41GB / 8.58GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 30%|███ | 8.87GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 30%|███ | 8.88GB / 29.6GB, 196MB/s
|
||
New Data Upload : 98%|█████████▊| 8.45GB / 8.65GB, 196MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 30%|███ | 8.90GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 30%|███ | 8.92GB / 29.6GB, 195MB/s
|
||
New Data Upload : 98%|█████████▊| 8.48GB / 8.65GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 30%|███ | 8.94GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 30%|███ | 8.96GB / 29.6GB, 193MB/s
|
||
New Data Upload : 98%|█████████▊| 8.52GB / 8.72GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 30%|███ | 8.97GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 30%|███ | 8.99GB / 29.6GB, 192MB/s
|
||
New Data Upload : 98%|█████████▊| 8.56GB / 8.72GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 30%|███ | 9.01GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 31%|███ | 9.03GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 8.59GB / 8.78GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 31%|███ | 9.03GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 31%|███ | 9.05GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 8.62GB / 8.78GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 31%|███ | 9.08GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 31%|███ | 9.10GB / 29.6GB, 192MB/s
|
||
New Data Upload : 98%|█████████▊| 8.66GB / 8.85GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 31%|███ | 9.12GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 31%|███ | 9.13GB / 29.6GB, 192MB/s
|
||
New Data Upload : 98%|█████████▊| 8.70GB / 8.85GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 31%|███ | 9.15GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 31%|███ | 9.16GB / 29.6GB, 192MB/s
|
||
New Data Upload : 98%|█████████▊| 8.73GB / 8.92GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 31%|███ | 9.18GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 31%|███ | 9.20GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 8.76GB / 8.92GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 31%|███ | 9.22GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 31%|███ | 9.24GB / 29.6GB, 190MB/s
|
||
New Data Upload : 98%|█████████▊| 8.80GB / 8.99GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 31%|███▏ | 9.25GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 31%|███▏ | 9.27GB / 29.6GB, 188MB/s
|
||
New Data Upload : 98%|█████████▊| 8.84GB / 9.05GB, 188MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 31%|███▏ | 9.28GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 31%|███▏ | 9.30GB / 29.6GB, 187MB/s
|
||
New Data Upload : 98%|█████████▊| 8.87GB / 9.05GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 32%|███▏ | 9.33GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 32%|███▏ | 9.34GB / 29.6GB, 187MB/s
|
||
New Data Upload : 98%|█████████▊| 8.91GB / 9.12GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 32%|███▏ | 9.37GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 32%|███▏ | 9.39GB / 29.6GB, 188MB/s
|
||
New Data Upload : 98%|█████████▊| 8.96GB / 9.12GB, 188MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 32%|███▏ | 9.41GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 32%|███▏ | 9.42GB / 29.6GB, 190MB/s
|
||
New Data Upload : 98%|█████████▊| 8.99GB / 9.19GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 32%|███▏ | 9.44GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 32%|███▏ | 9.46GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 9.03GB / 9.19GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 32%|███▏ | 9.46GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 32%|███▏ | 9.48GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 9.05GB / 9.25GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 32%|███▏ | 9.49GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 32%|███▏ | 9.51GB / 29.6GB, 190MB/s
|
||
New Data Upload : 98%|█████████▊| 9.07GB / 9.25GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 32%|███▏ | 9.53GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 32%|███▏ | 9.54GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 9.11GB / 9.32GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 32%|███▏ | 9.56GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 32%|███▏ | 9.58GB / 29.6GB, 190MB/s
|
||
New Data Upload : 97%|█████████▋| 9.15GB / 9.39GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 33%|███▎ | 9.61GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 33%|███▎ | 9.63GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 9.19GB / 9.39GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 33%|███▎ | 9.66GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 33%|███▎ | 9.68GB / 29.6GB, 193MB/s
|
||
New Data Upload : 98%|█████████▊| 9.25GB / 9.46GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 33%|███▎ | 9.71GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 33%|███▎ | 9.73GB / 29.6GB, 193MB/s
|
||
New Data Upload : 98%|█████████▊| 9.30GB / 9.46GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 33%|███▎ | 9.75GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 33%|███▎ | 9.77GB / 29.6GB, 192MB/s
|
||
New Data Upload : 98%|█████████▊| 9.33GB / 9.52GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 33%|███▎ | 9.77GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 33%|███▎ | 9.79GB / 29.6GB, 189MB/s
|
||
New Data Upload : 98%|█████████▊| 9.36GB / 9.52GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 33%|███▎ | 9.80GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 33%|███▎ | 9.82GB / 29.6GB, 188MB/s
|
||
New Data Upload : 98%|█████████▊| 9.39GB / 9.59GB, 188MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 33%|███▎ | 9.84GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 33%|███▎ | 9.86GB / 29.6GB, 188MB/s
|
||
New Data Upload : 98%|█████████▊| 9.43GB / 9.59GB, 188MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 33%|███▎ | 9.88GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 33%|███▎ | 9.90GB / 29.6GB, 187MB/s
|
||
New Data Upload : 98%|█████████▊| 9.46GB / 9.66GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 34%|███▎ | 9.92GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 34%|███▎ | 9.94GB / 29.6GB, 188MB/s
|
||
New Data Upload : 98%|█████████▊| 9.50GB / 9.66GB, 188MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 34%|███▎ | 9.95GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 34%|███▎ | 9.97GB / 29.6GB, 188MB/s
|
||
New Data Upload : 98%|█████████▊| 9.53GB / 9.72GB, 188MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 34%|███▎ | 9.96GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 34%|███▍ | 9.98GB / 29.6GB, 186MB/s
|
||
New Data Upload : 98%|█████████▊| 9.55GB / 9.72GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 34%|███▍ | 9.98GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 34%|███▍ | 10.0GB / 29.6GB, 184MB/s
|
||
New Data Upload : 98%|█████████▊| 9.57GB / 9.79GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 34%|███▍ | 10.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 34%|███▍ | 10.0GB / 29.6GB, 183MB/s
|
||
New Data Upload : 98%|█████████▊| 9.59GB / 9.79GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 34%|███▍ | 10.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 34%|███▍ | 10.1GB / 29.6GB, 183MB/s
|
||
New Data Upload : 98%|█████████▊| 9.63GB / 9.86GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 34%|███▍ | 10.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 34%|███▍ | 10.1GB / 29.6GB, 183MB/s
|
||
New Data Upload : 98%|█████████▊| 9.66GB / 9.86GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 34%|███▍ | 10.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 34%|███▍ | 10.1GB / 29.6GB, 184MB/s
|
||
New Data Upload : 98%|█████████▊| 9.71GB / 9.86GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 34%|███▍ | 10.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 34%|███▍ | 10.2GB / 29.6GB, 182MB/s
|
||
New Data Upload : 98%|█████████▊| 9.74GB / 9.92GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 35%|███▍ | 10.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 35%|███▍ | 10.2GB / 29.6GB, 181MB/s
|
||
New Data Upload : 98%|█████████▊| 9.78GB / 9.99GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 35%|███▍ | 10.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 35%|███▍ | 10.2GB / 29.6GB, 178MB/s
|
||
New Data Upload : 98%|█████████▊| 9.80GB / 9.99GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 35%|███▍ | 10.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 35%|███▍ | 10.3GB / 29.6GB, 175MB/s
|
||
New Data Upload : 98%|█████████▊| 9.82GB / 10.1GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 35%|███▍ | 10.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 35%|███▍ | 10.3GB / 29.6GB, 174MB/s
|
||
New Data Upload : 98%|█████████▊| 9.84GB / 10.1GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 35%|███▍ | 10.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 35%|███▍ | 10.3GB / 29.6GB, 175MB/s
|
||
New Data Upload : 98%|█████████▊| 9.88GB / 10.1GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 35%|███▍ | 10.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 35%|███▍ | 10.3GB / 29.6GB, 174MB/s
|
||
New Data Upload : 98%|█████████▊| 9.91GB / 10.1GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 35%|███▌ | 10.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 35%|███▌ | 10.4GB / 29.6GB, 175MB/s
|
||
New Data Upload : 98%|█████████▊| 9.96GB / 10.2GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 35%|███▌ | 10.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 35%|███▌ | 10.4GB / 29.6GB, 176MB/s
|
||
New Data Upload : 98%|█████████▊| 10.0GB / 10.2GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 35%|███▌ | 10.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 36%|███▌ | 10.5GB / 29.6GB, 178MB/s
|
||
New Data Upload : 98%|█████████▊| 10.1GB / 10.3GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 36%|███▌ | 10.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 36%|███▌ | 10.5GB / 29.6GB, 179MB/s
|
||
New Data Upload : 98%|█████████▊| 10.1GB / 10.3GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 36%|███▌ | 10.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 36%|███▌ | 10.6GB / 29.6GB, 178MB/s
|
||
New Data Upload : 98%|█████████▊| 10.1GB / 10.3GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 36%|███▌ | 10.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 36%|███▌ | 10.6GB / 29.6GB, 177MB/s
|
||
New Data Upload : 98%|█████████▊| 10.2GB / 10.4GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 36%|███▌ | 10.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 36%|███▌ | 10.6GB / 29.6GB, 176MB/s
|
||
New Data Upload : 98%|█████████▊| 10.2GB / 10.4GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 36%|███▌ | 10.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 36%|███▌ | 10.7GB / 29.6GB, 174MB/s
|
||
New Data Upload : 98%|█████████▊| 10.2GB / 10.5GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 36%|███▌ | 10.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 36%|███▌ | 10.7GB / 29.6GB, 176MB/s
|
||
New Data Upload : 98%|█████████▊| 10.3GB / 10.5GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 36%|███▋ | 10.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 36%|███▋ | 10.8GB / 29.6GB, 176MB/s
|
||
New Data Upload : 98%|█████████▊| 10.3GB / 10.5GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 37%|███▋ | 10.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 37%|███▋ | 10.8GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▊| 10.4GB / 10.5GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 37%|███▋ | 10.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 37%|███▋ | 10.8GB / 29.6GB, 179MB/s
|
||
New Data Upload : 98%|█████████▊| 10.4GB / 10.6GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 37%|███▋ | 10.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 37%|███▋ | 10.9GB / 29.6GB, 180MB/s
|
||
New Data Upload : 98%|█████████▊| 10.5GB / 10.7GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 37%|███▋ | 10.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 37%|███▋ | 10.9GB / 29.6GB, 179MB/s
|
||
New Data Upload : 98%|█████████▊| 10.5GB / 10.7GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 37%|███▋ | 10.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 37%|███▋ | 10.9GB / 29.6GB, 177MB/s
|
||
New Data Upload : 98%|█████████▊| 10.5GB / 10.7GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 37%|███▋ | 10.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 37%|███▋ | 11.0GB / 29.6GB, 176MB/s
|
||
New Data Upload : 98%|█████████▊| 10.5GB / 10.7GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 37%|███▋ | 11.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 37%|███▋ | 11.0GB / 29.6GB, 177MB/s
|
||
New Data Upload : 98%|█████████▊| 10.6GB / 10.8GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 37%|███▋ | 11.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 37%|███▋ | 11.0GB / 29.6GB, 177MB/s
|
||
New Data Upload : 98%|█████████▊| 10.6GB / 10.9GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 37%|███▋ | 11.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 37%|███▋ | 11.1GB / 29.6GB, 177MB/s
|
||
New Data Upload : 98%|█████████▊| 10.6GB / 10.9GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 38%|███▊ | 11.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 38%|███▊ | 11.1GB / 29.6GB, 179MB/s
|
||
New Data Upload : 98%|█████████▊| 10.7GB / 10.9GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 38%|███▊ | 11.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 38%|███▊ | 11.2GB / 29.6GB, 179MB/s
|
||
New Data Upload : 98%|█████████▊| 10.7GB / 10.9GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 38%|███▊ | 11.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 38%|███▊ | 11.2GB / 29.6GB, 178MB/s
|
||
New Data Upload : 98%|█████████▊| 10.8GB / 11.0GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 38%|███▊ | 11.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 38%|███▊ | 11.2GB / 29.6GB, 178MB/s
|
||
New Data Upload : 98%|█████████▊| 10.8GB / 11.1GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 38%|███▊ | 11.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 38%|███▊ | 11.3GB / 29.6GB, 178MB/s
|
||
New Data Upload : 98%|█████████▊| 10.8GB / 11.1GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 38%|███▊ | 11.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 38%|███▊ | 11.3GB / 29.6GB, 179MB/s
|
||
New Data Upload : 98%|█████████▊| 10.9GB / 11.1GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 38%|███▊ | 11.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 38%|███▊ | 11.3GB / 29.6GB, 180MB/s
|
||
New Data Upload : 98%|█████████▊| 10.9GB / 11.1GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 39%|███▊ | 11.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 39%|███▊ | 11.4GB / 29.6GB, 182MB/s
|
||
New Data Upload : 98%|█████████▊| 11.0GB / 11.2GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 39%|███▊ | 11.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 39%|███▊ | 11.4GB / 29.6GB, 183MB/s
|
||
New Data Upload : 98%|█████████▊| 11.0GB / 11.3GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 39%|███▉ | 11.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 39%|███▉ | 11.5GB / 29.6GB, 184MB/s
|
||
New Data Upload : 98%|█████████▊| 11.1GB / 11.3GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 39%|███▉ | 11.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 39%|███▉ | 11.5GB / 29.6GB, 183MB/s
|
||
New Data Upload : 98%|█████████▊| 11.1GB / 11.3GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 39%|███▉ | 11.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 39%|███▉ | 11.6GB / 29.6GB, 182MB/s
|
||
New Data Upload : 98%|█████████▊| 11.2GB / 11.4GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 39%|███▉ | 11.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 39%|███▉ | 11.6GB / 29.6GB, 184MB/s
|
||
New Data Upload : 98%|█████████▊| 11.2GB / 11.4GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 40%|███▉ | 11.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 40%|███▉ | 11.7GB / 29.6GB, 186MB/s
|
||
New Data Upload : 98%|█████████▊| 11.3GB / 11.5GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 40%|███▉ | 11.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 40%|███▉ | 11.7GB / 29.6GB, 186MB/s
|
||
New Data Upload : 98%|█████████▊| 11.3GB / 11.5GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 40%|███▉ | 11.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 40%|███▉ | 11.7GB / 29.6GB, 185MB/s
|
||
New Data Upload : 98%|█████████▊| 11.3GB / 11.5GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 40%|███▉ | 11.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 40%|███▉ | 11.8GB / 29.6GB, 186MB/s
|
||
New Data Upload : 98%|█████████▊| 11.4GB / 11.5GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 40%|███▉ | 11.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 40%|███▉ | 11.8GB / 29.6GB, 184MB/s
|
||
New Data Upload : 98%|█████████▊| 11.4GB / 11.6GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 40%|████ | 11.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 40%|████ | 11.8GB / 29.6GB, 184MB/s
|
||
New Data Upload : 98%|█████████▊| 11.4GB / 11.6GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 40%|████ | 11.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 40%|████ | 11.9GB / 29.6GB, 187MB/s
|
||
New Data Upload : 98%|█████████▊| 11.5GB / 11.7GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 40%|████ | 11.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 40%|████ | 11.9GB / 29.6GB, 188MB/s
|
||
New Data Upload : 98%|█████████▊| 11.5GB / 11.7GB, 188MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 40%|████ | 11.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 40%|████ | 12.0GB / 29.6GB, 190MB/s
|
||
New Data Upload : 98%|█████████▊| 11.5GB / 11.7GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 41%|████ | 12.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 41%|████ | 12.0GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 11.6GB / 11.8GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 41%|████ | 12.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 41%|████ | 12.1GB / 29.6GB, 192MB/s
|
||
New Data Upload : 98%|█████████▊| 11.6GB / 11.8GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 41%|████ | 12.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 41%|████ | 12.1GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 11.7GB / 11.9GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 41%|████ | 12.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 41%|████ | 12.1GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 11.7GB / 11.9GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 41%|████ | 12.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 41%|████ | 12.1GB / 29.6GB, 189MB/s
|
||
New Data Upload : 98%|█████████▊| 11.7GB / 11.9GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 41%|████ | 12.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 41%|████ | 12.2GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 11.7GB / 12.0GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 41%|████▏ | 12.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 41%|████▏ | 12.2GB / 29.6GB, 192MB/s
|
||
New Data Upload : 98%|█████████▊| 11.8GB / 12.0GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 41%|████▏ | 12.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 41%|████▏ | 12.3GB / 29.6GB, 194MB/s
|
||
New Data Upload : 98%|█████████▊| 11.8GB / 12.1GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 42%|████▏ | 12.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 42%|████▏ | 12.3GB / 29.6GB, 195MB/s
|
||
New Data Upload : 98%|█████████▊| 11.9GB / 12.1GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 42%|████▏ | 12.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 42%|████▏ | 12.4GB / 29.6GB, 197MB/s
|
||
New Data Upload : 98%|█████████▊| 11.9GB / 12.1GB, 197MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 42%|████▏ | 12.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 42%|████▏ | 12.4GB / 29.6GB, 198MB/s
|
||
New Data Upload : 99%|█████████▊| 12.0GB / 12.1GB, 198MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 42%|████▏ | 12.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 42%|████▏ | 12.5GB / 29.6GB, 198MB/s
|
||
New Data Upload : 99%|█████████▊| 12.0GB / 12.2GB, 198MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 42%|████▏ | 12.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 42%|████▏ | 12.5GB / 29.6GB, 197MB/s
|
||
New Data Upload : 99%|█████████▉| 12.1GB / 12.2GB, 197MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 42%|████▏ | 12.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 42%|████▏ | 12.5GB / 29.6GB, 195MB/s
|
||
New Data Upload : 99%|█████████▊| 12.1GB / 12.3GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 42%|████▏ | 12.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 42%|████▏ | 12.6GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▉| 12.1GB / 12.3GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 43%|████▎ | 12.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 43%|████▎ | 12.6GB / 29.6GB, 195MB/s
|
||
New Data Upload : 99%|█████████▊| 12.2GB / 12.3GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 43%|████▎ | 12.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 43%|████▎ | 12.6GB / 29.6GB, 196MB/s
|
||
New Data Upload : 98%|█████████▊| 12.2GB / 12.4GB, 196MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 43%|████▎ | 12.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 43%|████▎ | 12.7GB / 29.6GB, 195MB/s
|
||
New Data Upload : 99%|█████████▊| 12.2GB / 12.4GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 43%|████▎ | 12.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 43%|████▎ | 12.7GB / 29.6GB, 194MB/s
|
||
New Data Upload : 98%|█████████▊| 12.3GB / 12.5GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 43%|████▎ | 12.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 43%|████▎ | 12.7GB / 29.6GB, 195MB/s
|
||
New Data Upload : 99%|█████████▊| 12.3GB / 12.5GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 43%|████▎ | 12.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 43%|████▎ | 12.8GB / 29.6GB, 194MB/s
|
||
New Data Upload : 98%|█████████▊| 12.3GB / 12.5GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 43%|████▎ | 12.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 43%|████▎ | 12.8GB / 29.6GB, 193MB/s
|
||
New Data Upload : 98%|█████████▊| 12.4GB / 12.6GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 44%|████▎ | 12.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 44%|████▎ | 12.9GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▊| 12.4GB / 12.6GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 44%|████▎ | 12.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 44%|████▎ | 12.9GB / 29.6GB, 195MB/s
|
||
New Data Upload : 98%|█████████▊| 12.5GB / 12.7GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 44%|████▍ | 12.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 44%|████▍ | 13.0GB / 29.6GB, 197MB/s
|
||
New Data Upload : 98%|█████████▊| 12.5GB / 12.7GB, 197MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 44%|████▍ | 13.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 44%|████▍ | 13.0GB / 29.6GB, 199MB/s
|
||
New Data Upload : 99%|█████████▊| 12.6GB / 12.7GB, 199MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 44%|████▍ | 13.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 44%|████▍ | 13.0GB / 29.6GB, 198MB/s
|
||
New Data Upload : 98%|█████████▊| 12.6GB / 12.8GB, 198MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 44%|████▍ | 13.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 44%|████▍ | 13.1GB / 29.6GB, 198MB/s
|
||
New Data Upload : 99%|█████████▊| 12.6GB / 12.8GB, 198MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 44%|████▍ | 13.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 44%|████▍ | 13.1GB / 29.6GB, 198MB/s
|
||
New Data Upload : 98%|█████████▊| 12.7GB / 12.9GB, 198MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 44%|████▍ | 13.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 44%|████▍ | 13.1GB / 29.6GB, 197MB/s
|
||
New Data Upload : 99%|█████████▊| 12.7GB / 12.9GB, 197MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 44%|████▍ | 13.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 45%|████▍ | 13.2GB / 29.6GB, 195MB/s
|
||
New Data Upload : 98%|█████████▊| 12.7GB / 12.9GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 45%|████▍ | 13.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 45%|████▍ | 13.2GB / 29.6GB, 196MB/s
|
||
New Data Upload : 99%|█████████▊| 12.8GB / 12.9GB, 196MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 45%|████▍ | 13.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 45%|████▍ | 13.2GB / 29.6GB, 197MB/s
|
||
New Data Upload : 98%|█████████▊| 12.8GB / 13.0GB, 197MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 45%|████▍ | 13.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 45%|████▍ | 13.3GB / 29.6GB, 197MB/s
|
||
New Data Upload : 98%|█████████▊| 12.8GB / 13.1GB, 197MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 45%|████▌ | 13.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 45%|████▌ | 13.3GB / 29.6GB, 196MB/s
|
||
New Data Upload : 98%|█████████▊| 12.9GB / 13.1GB, 196MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 45%|████▌ | 13.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 45%|████▌ | 13.3GB / 29.6GB, 196MB/s
|
||
New Data Upload : 98%|█████████▊| 12.9GB / 13.1GB, 196MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 45%|████▌ | 13.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 45%|████▌ | 13.4GB / 29.6GB, 194MB/s
|
||
New Data Upload : 98%|█████████▊| 12.9GB / 13.1GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 45%|████▌ | 13.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 45%|████▌ | 13.4GB / 29.6GB, 192MB/s
|
||
New Data Upload : 98%|█████████▊| 13.0GB / 13.2GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 45%|████▌ | 13.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 45%|████▌ | 13.4GB / 29.6GB, 190MB/s
|
||
New Data Upload : 98%|█████████▊| 13.0GB / 13.2GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 46%|████▌ | 13.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 46%|████▌ | 13.5GB / 29.6GB, 189MB/s
|
||
New Data Upload : 98%|█████████▊| 13.0GB / 13.3GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 46%|████▌ | 13.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 46%|████▌ | 13.5GB / 29.6GB, 188MB/s
|
||
New Data Upload : 98%|█████████▊| 13.1GB / 13.3GB, 188MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 46%|████▌ | 13.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 46%|████▌ | 13.5GB / 29.6GB, 186MB/s
|
||
New Data Upload : 98%|█████████▊| 13.1GB / 13.3GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 46%|████▌ | 13.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 46%|████▌ | 13.6GB / 29.6GB, 185MB/s
|
||
New Data Upload : 98%|█████████▊| 13.1GB / 13.4GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 46%|████▌ | 13.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 46%|████▌ | 13.6GB / 29.6GB, 186MB/s
|
||
New Data Upload : 98%|█████████▊| 13.2GB / 13.4GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 46%|████▌ | 13.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 46%|████▌ | 13.6GB / 29.6GB, 186MB/s
|
||
New Data Upload : 98%|█████████▊| 13.2GB / 13.5GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 46%|████▋ | 13.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 46%|████▋ | 13.7GB / 29.6GB, 186MB/s
|
||
New Data Upload : 98%|█████████▊| 13.3GB / 13.5GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 46%|████▋ | 13.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 46%|████▋ | 13.7GB / 29.6GB, 188MB/s
|
||
New Data Upload : 98%|█████████▊| 13.3GB / 13.5GB, 188MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 47%|████▋ | 13.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 47%|████▋ | 13.8GB / 29.6GB, 189MB/s
|
||
New Data Upload : 98%|█████████▊| 13.3GB / 13.5GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 47%|████▋ | 13.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 47%|████▋ | 13.8GB / 29.6GB, 189MB/s
|
||
New Data Upload : 98%|█████████▊| 13.4GB / 13.6GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 47%|████▋ | 13.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 47%|████▋ | 13.9GB / 29.6GB, 190MB/s
|
||
New Data Upload : 99%|█████████▊| 13.4GB / 13.6GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 47%|████▋ | 13.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 47%|████▋ | 13.9GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 13.5GB / 13.7GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 47%|████▋ | 13.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 47%|████▋ | 14.0GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 13.5GB / 13.7GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 47%|████▋ | 14.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 47%|████▋ | 14.0GB / 29.6GB, 191MB/s
|
||
New Data Upload : 99%|█████████▊| 13.6GB / 13.7GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 47%|████▋ | 14.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 47%|████▋ | 14.0GB / 29.6GB, 189MB/s
|
||
New Data Upload : 98%|█████████▊| 13.6GB / 13.8GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 47%|████▋ | 14.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 47%|████▋ | 14.0GB / 29.6GB, 189MB/s
|
||
New Data Upload : 99%|█████████▊| 13.6GB / 13.8GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 48%|████▊ | 14.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 48%|████▊ | 14.1GB / 29.6GB, 190MB/s
|
||
New Data Upload : 98%|█████████▊| 13.6GB / 13.9GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 48%|████▊ | 14.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 48%|████▊ | 14.1GB / 29.6GB, 191MB/s
|
||
New Data Upload : 98%|█████████▊| 13.7GB / 13.9GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 48%|████▊ | 14.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 48%|████▊ | 14.2GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▊| 13.7GB / 13.9GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 48%|████▊ | 14.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 48%|████▊ | 14.2GB / 29.6GB, 194MB/s
|
||
New Data Upload : 98%|█████████▊| 13.8GB / 14.0GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 48%|████▊ | 14.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 48%|████▊ | 14.3GB / 29.6GB, 194MB/s
|
||
New Data Upload : 98%|█████████▊| 13.8GB / 14.1GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 48%|████▊ | 14.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 48%|████▊ | 14.3GB / 29.6GB, 192MB/s
|
||
New Data Upload : 99%|█████████▊| 13.9GB / 14.1GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 48%|████▊ | 14.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 48%|████▊ | 14.3GB / 29.6GB, 189MB/s
|
||
New Data Upload : 98%|█████████▊| 13.9GB / 14.2GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 49%|████▊ | 14.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 49%|████▊ | 14.4GB / 29.6GB, 186MB/s
|
||
New Data Upload : 98%|█████████▊| 13.9GB / 14.2GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 49%|████▊ | 14.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 49%|████▊ | 14.4GB / 29.6GB, 185MB/s
|
||
New Data Upload : 98%|█████████▊| 14.0GB / 14.2GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 49%|████▉ | 14.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 49%|████▉ | 14.4GB / 29.6GB, 187MB/s
|
||
New Data Upload : 98%|█████████▊| 14.0GB / 14.2GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 49%|████▉ | 14.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 49%|████▉ | 14.5GB / 29.6GB, 189MB/s
|
||
New Data Upload : 98%|█████████▊| 14.0GB / 14.3GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 49%|████▉ | 14.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 49%|████▉ | 14.5GB / 29.6GB, 189MB/s
|
||
New Data Upload : 99%|█████████▊| 14.1GB / 14.3GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 49%|████▉ | 14.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 49%|████▉ | 14.6GB / 29.6GB, 191MB/s
|
||
New Data Upload : 99%|█████████▊| 14.1GB / 14.4GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 49%|████▉ | 14.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 49%|████▉ | 14.6GB / 29.6GB, 192MB/s
|
||
New Data Upload : 99%|█████████▉| 14.2GB / 14.4GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 50%|████▉ | 14.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 50%|████▉ | 14.6GB / 29.6GB, 192MB/s
|
||
New Data Upload : 99%|█████████▊| 14.2GB / 14.4GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 50%|████▉ | 14.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 50%|████▉ | 14.7GB / 29.6GB, 190MB/s
|
||
New Data Upload : 99%|█████████▉| 14.2GB / 14.4GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 50%|████▉ | 14.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 50%|████▉ | 14.7GB / 29.6GB, 189MB/s
|
||
New Data Upload : 99%|█████████▊| 14.3GB / 14.5GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 50%|████▉ | 14.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 50%|████▉ | 14.7GB / 29.6GB, 188MB/s
|
||
New Data Upload : 99%|█████████▉| 14.3GB / 14.5GB, 188MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 50%|████▉ | 14.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 50%|████▉ | 14.8GB / 29.6GB, 187MB/s
|
||
New Data Upload : 99%|█████████▊| 14.3GB / 14.6GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 50%|█████ | 14.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 50%|█████ | 14.8GB / 29.6GB, 186MB/s
|
||
New Data Upload : 99%|█████████▉| 14.4GB / 14.6GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 50%|█████ | 14.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 50%|█████ | 14.8GB / 29.6GB, 186MB/s
|
||
New Data Upload : 99%|█████████▊| 14.4GB / 14.6GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 50%|█████ | 14.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 50%|█████ | 14.9GB / 29.6GB, 187MB/s
|
||
New Data Upload : 98%|█████████▊| 14.5GB / 14.7GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 51%|█████ | 14.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 51%|█████ | 14.9GB / 29.6GB, 187MB/s
|
||
New Data Upload : 99%|█████████▊| 14.5GB / 14.7GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 51%|█████ | 14.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 51%|█████ | 15.0GB / 29.6GB, 187MB/s
|
||
New Data Upload : 99%|█████████▊| 14.5GB / 14.8GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 51%|█████ | 15.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 51%|█████ | 15.0GB / 29.6GB, 187MB/s
|
||
New Data Upload : 99%|█████████▉| 14.6GB / 14.8GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 51%|█████ | 15.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 51%|█████ | 15.0GB / 29.6GB, 187MB/s
|
||
New Data Upload : 99%|█████████▊| 14.6GB / 14.8GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 51%|█████ | 15.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 51%|█████ | 15.1GB / 29.6GB, 189MB/s
|
||
New Data Upload : 99%|█████████▉| 14.7GB / 14.8GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 51%|█████ | 15.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 51%|█████ | 15.1GB / 29.6GB, 187MB/s
|
||
New Data Upload : 99%|█████████▊| 14.7GB / 14.9GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 51%|█████ | 15.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 51%|█████ | 15.1GB / 29.6GB, 186MB/s
|
||
New Data Upload : 98%|█████████▊| 14.7GB / 15.0GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 51%|█████▏ | 15.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 51%|█████▏ | 15.2GB / 29.6GB, 184MB/s
|
||
New Data Upload : 98%|█████████▊| 14.7GB / 15.0GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 51%|█████▏ | 15.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 51%|█████▏ | 15.2GB / 29.6GB, 184MB/s
|
||
New Data Upload : 98%|█████████▊| 14.8GB / 15.0GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 51%|█████▏ | 15.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 51%|█████▏ | 15.2GB / 29.6GB, 183MB/s
|
||
New Data Upload : 98%|█████████▊| 14.8GB / 15.0GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 52%|█████▏ | 15.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 52%|█████▏ | 15.2GB / 29.6GB, 183MB/s
|
||
New Data Upload : 98%|█████████▊| 14.8GB / 15.1GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 52%|█████▏ | 15.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 52%|█████▏ | 15.3GB / 29.6GB, 185MB/s
|
||
New Data Upload : 98%|█████████▊| 14.9GB / 15.2GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 52%|█████▏ | 15.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 52%|█████▏ | 15.3GB / 29.6GB, 186MB/s
|
||
New Data Upload : 98%|█████████▊| 14.9GB / 15.2GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 52%|█████▏ | 15.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 52%|█████▏ | 15.4GB / 29.6GB, 189MB/s
|
||
New Data Upload : 98%|█████████▊| 15.0GB / 15.2GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 52%|█████▏ | 15.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 52%|█████▏ | 15.4GB / 29.6GB, 189MB/s
|
||
New Data Upload : 99%|█████████▊| 15.0GB / 15.2GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 52%|█████▏ | 15.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 52%|█████▏ | 15.5GB / 29.6GB, 191MB/s
|
||
New Data Upload : 99%|█████████▉| 15.0GB / 15.2GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 52%|█████▏ | 15.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 52%|█████▏ | 15.5GB / 29.6GB, 190MB/s
|
||
New Data Upload : 99%|█████████▊| 15.1GB / 15.3GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 53%|█████▎ | 15.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 53%|█████▎ | 15.6GB / 29.6GB, 190MB/s
|
||
New Data Upload : 98%|█████████▊| 15.1GB / 15.4GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 53%|█████▎ | 15.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 53%|█████▎ | 15.6GB / 29.6GB, 190MB/s
|
||
New Data Upload : 99%|█████████▊| 15.1GB / 15.4GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 53%|█████▎ | 15.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 53%|█████▎ | 15.6GB / 29.6GB, 189MB/s
|
||
New Data Upload : 99%|█████████▉| 15.2GB / 15.4GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 53%|█████▎ | 15.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 53%|█████▎ | 15.7GB / 29.6GB, 190MB/s
|
||
New Data Upload : 99%|█████████▉| 15.2GB / 15.4GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 53%|█████▎ | 15.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 53%|█████▎ | 15.7GB / 29.6GB, 191MB/s
|
||
New Data Upload : 99%|█████████▉| 15.3GB / 15.4GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 53%|█████▎ | 15.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 53%|█████▎ | 15.8GB / 29.6GB, 191MB/s
|
||
New Data Upload : 99%|█████████▉| 15.3GB / 15.5GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 53%|█████▎ | 15.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 53%|█████▎ | 15.8GB / 29.6GB, 191MB/s
|
||
New Data Upload : 99%|█████████▉| 15.4GB / 15.6GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 54%|█████▎ | 15.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 54%|█████▎ | 15.8GB / 29.6GB, 188MB/s
|
||
New Data Upload : 99%|█████████▉| 15.4GB / 15.6GB, 188MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 54%|█████▎ | 15.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 54%|█████▎ | 15.8GB / 29.6GB, 185MB/s
|
||
New Data Upload : 99%|█████████▉| 15.4GB / 15.6GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 54%|█████▎ | 15.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 54%|█████▎ | 15.9GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 15.4GB / 15.6GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 54%|█████▍ | 15.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 54%|█████▍ | 15.9GB / 29.6GB, 185MB/s
|
||
New Data Upload : 99%|█████████▉| 15.5GB / 15.6GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 54%|█████▍ | 15.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 54%|█████▍ | 15.9GB / 29.6GB, 185MB/s
|
||
New Data Upload : 99%|█████████▉| 15.5GB / 15.7GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 54%|█████▍ | 15.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 54%|█████▍ | 16.0GB / 29.6GB, 184MB/s
|
||
New Data Upload : 99%|█████████▉| 15.5GB / 15.7GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 54%|█████▍ | 16.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 54%|█████▍ | 16.0GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▊| 15.6GB / 15.8GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 54%|█████▍ | 16.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 54%|█████▍ | 16.0GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 15.6GB / 15.8GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 54%|█████▍ | 16.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 54%|█████▍ | 16.1GB / 29.6GB, 180MB/s
|
||
New Data Upload : 99%|█████████▉| 15.6GB / 15.8GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 54%|█████▍ | 16.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 54%|█████▍ | 16.1GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 15.7GB / 15.8GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 55%|█████▍ | 16.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 55%|█████▍ | 16.1GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▊| 15.7GB / 15.9GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 55%|█████▍ | 16.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 55%|█████▍ | 16.2GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▊| 15.7GB / 16.0GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 55%|█████▍ | 16.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 55%|█████▍ | 16.2GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▊| 15.8GB / 16.0GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 55%|█████▍ | 16.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 55%|█████▍ | 16.2GB / 29.6GB, 178MB/s
|
||
New Data Upload : 98%|█████████▊| 15.8GB / 16.0GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 55%|█████▍ | 16.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 55%|█████▍ | 16.3GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▊| 15.8GB / 16.0GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 55%|█████▌ | 16.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 55%|█████▌ | 16.3GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 15.9GB / 16.0GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 55%|█████▌ | 16.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 55%|█████▌ | 16.3GB / 29.6GB, 176MB/s
|
||
New Data Upload : 99%|█████████▊| 15.9GB / 16.1GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 55%|█████▌ | 16.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 55%|█████▌ | 16.4GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 15.9GB / 16.1GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 55%|█████▌ | 16.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 55%|█████▌ | 16.4GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▊| 16.0GB / 16.2GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 56%|█████▌ | 16.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 56%|█████▌ | 16.4GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 16.0GB / 16.2GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 56%|█████▌ | 16.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 56%|█████▌ | 16.4GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▊| 16.0GB / 16.2GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 56%|█████▌ | 16.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 56%|█████▌ | 16.5GB / 29.6GB, 173MB/s
|
||
New Data Upload : 98%|█████████▊| 16.0GB / 16.3GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 56%|█████▌ | 16.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 56%|█████▌ | 16.5GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▊| 16.1GB / 16.3GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 56%|█████▌ | 16.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 56%|█████▌ | 16.5GB / 29.6GB, 173MB/s
|
||
New Data Upload : 98%|█████████▊| 16.1GB / 16.4GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 56%|█████▌ | 16.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 56%|█████▌ | 16.6GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▊| 16.1GB / 16.4GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 56%|█████▌ | 16.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 56%|█████▌ | 16.6GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▊| 16.2GB / 16.4GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 56%|█████▋ | 16.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 56%|█████▋ | 16.7GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 16.2GB / 16.4GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 56%|█████▋ | 16.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 56%|█████▋ | 16.7GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▊| 16.3GB / 16.5GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 57%|█████▋ | 16.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 57%|█████▋ | 16.7GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 16.3GB / 16.5GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 57%|█████▋ | 16.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 57%|█████▋ | 16.8GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▊| 16.3GB / 16.6GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 57%|█████▋ | 16.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 57%|█████▋ | 16.8GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▊| 16.4GB / 16.6GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 57%|█████▋ | 16.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 57%|█████▋ | 16.9GB / 29.6GB, 176MB/s
|
||
New Data Upload : 99%|█████████▉| 16.5GB / 16.6GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 57%|█████▋ | 16.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 57%|█████▋ | 16.9GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 16.5GB / 16.7GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 57%|█████▋ | 16.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 57%|█████▋ | 17.0GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 16.5GB / 16.7GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 57%|█████▋ | 17.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 57%|█████▋ | 17.0GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▊| 16.6GB / 16.8GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 58%|█████▊ | 17.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 58%|█████▊ | 17.0GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 16.6GB / 16.8GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 58%|█████▊ | 17.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 58%|█████▊ | 17.0GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▉| 16.6GB / 16.8GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 58%|█████▊ | 17.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 58%|█████▊ | 17.1GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▉| 16.6GB / 16.8GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 58%|█████▊ | 17.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 58%|█████▊ | 17.1GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 16.7GB / 16.8GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 58%|█████▊ | 17.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 58%|█████▊ | 17.1GB / 29.6GB, 176MB/s
|
||
New Data Upload : 99%|█████████▉| 16.7GB / 16.9GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 58%|█████▊ | 17.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 58%|█████▊ | 17.2GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 16.7GB / 16.9GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 58%|█████▊ | 17.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 58%|█████▊ | 17.2GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 16.8GB / 17.0GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 58%|█████▊ | 17.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 58%|█████▊ | 17.2GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 16.8GB / 17.0GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 58%|█████▊ | 17.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 58%|█████▊ | 17.3GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 16.8GB / 17.0GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 59%|█████▊ | 17.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 59%|█████▊ | 17.3GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▊| 16.9GB / 17.1GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 59%|█████▊ | 17.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 59%|█████▊ | 17.3GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 16.9GB / 17.1GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 59%|█████▉ | 17.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 59%|█████▉ | 17.4GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▊| 16.9GB / 17.2GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 59%|█████▉ | 17.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 59%|█████▉ | 17.4GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 17.0GB / 17.2GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 59%|█████▉ | 17.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 59%|█████▉ | 17.5GB / 29.6GB, 170MB/s
|
||
New Data Upload : 99%|█████████▊| 17.0GB / 17.2GB, 170MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 59%|█████▉ | 17.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 59%|█████▉ | 17.5GB / 29.6GB, 168MB/s
|
||
New Data Upload : 99%|█████████▉| 17.0GB / 17.2GB, 168MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 59%|█████▉ | 17.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 59%|█████▉ | 17.5GB / 29.6GB, 167MB/s
|
||
New Data Upload : 99%|█████████▊| 17.1GB / 17.3GB, 167MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 59%|█████▉ | 17.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 59%|█████▉ | 17.5GB / 29.6GB, 169MB/s
|
||
New Data Upload : 99%|█████████▉| 17.1GB / 17.3GB, 169MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 59%|█████▉ | 17.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 59%|█████▉ | 17.6GB / 29.6GB, 170MB/s
|
||
New Data Upload : 99%|█████████▊| 17.1GB / 17.4GB, 170MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 60%|█████▉ | 17.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 60%|█████▉ | 17.6GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 17.2GB / 17.4GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 60%|█████▉ | 17.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 60%|█████▉ | 17.7GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 17.2GB / 17.4GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 60%|█████▉ | 17.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 60%|█████▉ | 17.7GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▊| 17.3GB / 17.5GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 60%|█████▉ | 17.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 60%|█████▉ | 17.7GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▊| 17.3GB / 17.5GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 60%|██████ | 17.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 60%|██████ | 17.7GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▊| 17.3GB / 17.6GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 60%|██████ | 17.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 60%|██████ | 17.8GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▊| 17.3GB / 17.6GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 60%|██████ | 17.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 60%|██████ | 17.8GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▊| 17.4GB / 17.6GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 60%|██████ | 17.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 60%|██████ | 17.9GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 17.4GB / 17.6GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 61%|██████ | 17.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 61%|██████ | 17.9GB / 29.6GB, 175MB/s
|
||
New Data Upload : 99%|█████████▊| 17.5GB / 17.7GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 61%|██████ | 17.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 61%|██████ | 17.9GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 17.5GB / 17.7GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 61%|██████ | 18.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 61%|██████ | 18.0GB / 29.6GB, 176MB/s
|
||
New Data Upload : 99%|█████████▊| 17.5GB / 17.8GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 61%|██████ | 18.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 61%|██████ | 18.0GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 17.6GB / 17.8GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 61%|██████ | 18.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 61%|██████ | 18.1GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 17.6GB / 17.8GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 61%|██████ | 18.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 61%|██████ | 18.1GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▊| 17.7GB / 17.9GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 61%|██████▏ | 18.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 61%|██████▏ | 18.1GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 17.7GB / 17.9GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 61%|██████▏ | 18.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 61%|██████▏ | 18.2GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▊| 17.7GB / 18.0GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 62%|██████▏ | 18.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 62%|██████▏ | 18.2GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▊| 17.8GB / 18.0GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 62%|██████▏ | 18.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 62%|██████▏ | 18.2GB / 29.6GB, 180MB/s
|
||
New Data Upload : 99%|█████████▉| 17.8GB / 18.0GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 62%|██████▏ | 18.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 62%|██████▏ | 18.3GB / 29.6GB, 180MB/s
|
||
New Data Upload : 99%|█████████▊| 17.8GB / 18.1GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 62%|██████▏ | 18.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 62%|██████▏ | 18.3GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▉| 17.9GB / 18.1GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 62%|██████▏ | 18.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 62%|██████▏ | 18.4GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▊| 17.9GB / 18.2GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 62%|██████▏ | 18.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 62%|██████▏ | 18.4GB / 29.6GB, 182MB/s
|
||
New Data Upload : 98%|█████████▊| 18.0GB / 18.2GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 62%|██████▏ | 18.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 62%|██████▏ | 18.4GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▊| 18.0GB / 18.2GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 63%|██████▎ | 18.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 63%|██████▎ | 18.5GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 18.1GB / 18.2GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 63%|██████▎ | 18.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 63%|██████▎ | 18.5GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 18.1GB / 18.3GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 63%|██████▎ | 18.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 63%|██████▎ | 18.6GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▊| 18.1GB / 18.4GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 63%|██████▎ | 18.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 63%|██████▎ | 18.6GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 18.2GB / 18.4GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 63%|██████▎ | 18.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 63%|██████▎ | 18.6GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▊| 18.2GB / 18.4GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 63%|██████▎ | 18.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 63%|██████▎ | 18.7GB / 29.6GB, 179MB/s
|
||
New Data Upload : 98%|█████████▊| 18.2GB / 18.5GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 63%|██████▎ | 18.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 63%|██████▎ | 18.7GB / 29.6GB, 176MB/s
|
||
New Data Upload : 99%|█████████▊| 18.3GB / 18.5GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 63%|██████▎ | 18.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 63%|██████▎ | 18.7GB / 29.6GB, 177MB/s
|
||
New Data Upload : 98%|█████████▊| 18.3GB / 18.6GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 64%|██████▎ | 18.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 64%|██████▎ | 18.8GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▊| 18.3GB / 18.6GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 64%|██████▎ | 18.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 64%|██████▎ | 18.8GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▊| 18.4GB / 18.6GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 64%|██████▍ | 18.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 64%|██████▍ | 18.9GB / 29.6GB, 182MB/s
|
||
New Data Upload : 98%|█████████▊| 18.4GB / 18.7GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 64%|██████▍ | 18.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 64%|██████▍ | 18.9GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▊| 18.5GB / 18.7GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 64%|██████▍ | 18.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 64%|██████▍ | 18.9GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▊| 18.5GB / 18.8GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 64%|██████▍ | 19.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 64%|██████▍ | 19.0GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▊| 18.5GB / 18.8GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 64%|██████▍ | 19.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 64%|██████▍ | 19.0GB / 29.6GB, 184MB/s
|
||
New Data Upload : 99%|█████████▊| 18.6GB / 18.8GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 64%|██████▍ | 19.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 64%|██████▍ | 19.1GB / 29.6GB, 185MB/s
|
||
New Data Upload : 99%|█████████▉| 18.6GB / 18.8GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 65%|██████▍ | 19.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 65%|██████▍ | 19.1GB / 29.6GB, 186MB/s
|
||
New Data Upload : 99%|█████████▊| 18.7GB / 18.9GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 65%|██████▍ | 19.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 65%|██████▍ | 19.1GB / 29.6GB, 187MB/s
|
||
New Data Upload : 99%|█████████▉| 18.7GB / 18.9GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 65%|██████▍ | 19.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 65%|██████▍ | 19.2GB / 29.6GB, 189MB/s
|
||
New Data Upload : 99%|█████████▉| 18.8GB / 19.0GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 65%|██████▌ | 19.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 65%|██████▌ | 19.2GB / 29.6GB, 189MB/s
|
||
New Data Upload : 99%|█████████▊| 18.8GB / 19.0GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 65%|██████▌ | 19.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 65%|██████▌ | 19.3GB / 29.6GB, 189MB/s
|
||
New Data Upload : 99%|█████████▉| 18.8GB / 19.0GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 65%|██████▌ | 19.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 65%|██████▌ | 19.3GB / 29.6GB, 189MB/s
|
||
New Data Upload : 99%|█████████▊| 18.9GB / 19.1GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 65%|██████▌ | 19.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 65%|██████▌ | 19.3GB / 29.6GB, 189MB/s
|
||
New Data Upload : 99%|█████████▊| 18.9GB / 19.2GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 66%|██████▌ | 19.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 66%|██████▌ | 19.4GB / 29.6GB, 190MB/s
|
||
New Data Upload : 99%|█████████▉| 19.0GB / 19.2GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 66%|██████▌ | 19.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 66%|██████▌ | 19.4GB / 29.6GB, 192MB/s
|
||
New Data Upload : 99%|█████████▊| 19.0GB / 19.2GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 66%|██████▌ | 19.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 66%|██████▌ | 19.5GB / 29.6GB, 192MB/s
|
||
New Data Upload : 99%|█████████▉| 19.0GB / 19.2GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 66%|██████▌ | 19.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 66%|██████▌ | 19.5GB / 29.6GB, 193MB/s
|
||
New Data Upload : 99%|█████████▉| 19.1GB / 19.3GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 66%|██████▌ | 19.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 66%|██████▌ | 19.6GB / 29.6GB, 193MB/s
|
||
New Data Upload : 99%|█████████▉| 19.1GB / 19.3GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 66%|██████▋ | 19.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 66%|██████▋ | 19.6GB / 29.6GB, 193MB/s
|
||
New Data Upload : 99%|█████████▉| 19.2GB / 19.4GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 66%|██████▋ | 19.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 66%|██████▋ | 19.6GB / 29.6GB, 193MB/s
|
||
New Data Upload : 99%|█████████▊| 19.2GB / 19.4GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 67%|██████▋ | 19.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 67%|██████▋ | 19.7GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▉| 19.2GB / 19.4GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 67%|██████▋ | 19.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 67%|██████▋ | 19.7GB / 29.6GB, 195MB/s
|
||
New Data Upload : 99%|█████████▉| 19.3GB / 19.5GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 67%|██████▋ | 19.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 67%|██████▋ | 19.7GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▉| 19.3GB / 19.5GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 67%|██████▋ | 19.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 67%|██████▋ | 19.8GB / 29.6GB, 195MB/s
|
||
New Data Upload : 99%|█████████▊| 19.3GB / 19.6GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 67%|██████▋ | 19.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 67%|██████▋ | 19.8GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▉| 19.4GB / 19.6GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 67%|██████▋ | 19.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 67%|██████▋ | 19.8GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▉| 19.4GB / 19.6GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 67%|██████▋ | 19.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 67%|██████▋ | 19.9GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▊| 19.4GB / 19.7GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 67%|██████▋ | 19.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 67%|██████▋ | 19.9GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▉| 19.5GB / 19.7GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 68%|██████▊ | 19.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 68%|██████▊ | 20.0GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▊| 19.5GB / 19.8GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 68%|██████▊ | 20.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 68%|██████▊ | 20.0GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▉| 19.6GB / 19.8GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 68%|██████▊ | 20.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 68%|██████▊ | 20.0GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▊| 19.6GB / 19.9GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 68%|██████▊ | 20.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 68%|██████▊ | 20.1GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▊| 19.6GB / 19.9GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 68%|██████▊ | 20.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 68%|██████▊ | 20.1GB / 29.6GB, 195MB/s
|
||
New Data Upload : 99%|█████████▉| 19.7GB / 19.9GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 68%|██████▊ | 20.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 68%|██████▊ | 20.2GB / 29.6GB, 195MB/s
|
||
New Data Upload : 99%|█████████▊| 19.7GB / 20.0GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 68%|██████▊ | 20.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 68%|██████▊ | 20.2GB / 29.6GB, 195MB/s
|
||
New Data Upload : 99%|█████████▉| 19.8GB / 20.0GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 68%|██████▊ | 20.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 68%|██████▊ | 20.2GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▊| 19.8GB / 20.1GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 69%|██████▊ | 20.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 69%|██████▊ | 20.3GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▊| 19.8GB / 20.1GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 69%|██████▊ | 20.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 69%|██████▊ | 20.3GB / 29.6GB, 194MB/s
|
||
New Data Upload : 99%|█████████▊| 19.9GB / 20.1GB, 194MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 69%|██████▉ | 20.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 69%|██████▉ | 20.3GB / 29.6GB, 195MB/s
|
||
New Data Upload : 99%|█████████▊| 19.9GB / 20.2GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 69%|██████▉ | 20.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 69%|██████▉ | 20.4GB / 29.6GB, 196MB/s
|
||
New Data Upload : 99%|█████████▉| 20.0GB / 20.2GB, 196MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 69%|██████▉ | 20.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 69%|██████▉ | 20.4GB / 29.6GB, 196MB/s
|
||
New Data Upload : 99%|█████████▉| 20.0GB / 20.3GB, 196MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 69%|██████▉ | 20.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 69%|██████▉ | 20.5GB / 29.6GB, 196MB/s
|
||
New Data Upload : 99%|█████████▉| 20.1GB / 20.3GB, 196MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 69%|██████▉ | 20.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 69%|██████▉ | 20.5GB / 29.6GB, 197MB/s
|
||
New Data Upload : 99%|█████████▉| 20.1GB / 20.3GB, 197MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 70%|██████▉ | 20.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 70%|██████▉ | 20.6GB / 29.6GB, 199MB/s
|
||
New Data Upload : 99%|█████████▉| 20.2GB / 20.3GB, 199MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 70%|██████▉ | 20.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 70%|██████▉ | 20.6GB / 29.6GB, 200MB/s
|
||
New Data Upload : 99%|█████████▉| 20.2GB / 20.4GB, 200MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 70%|██████▉ | 20.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 70%|██████▉ | 20.6GB / 29.6GB, 198MB/s
|
||
New Data Upload : 99%|█████████▉| 20.2GB / 20.4GB, 198MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 70%|██████▉ | 20.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 70%|██████▉ | 20.7GB / 29.6GB, 198MB/s
|
||
New Data Upload : 99%|█████████▉| 20.2GB / 20.5GB, 198MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 70%|███████ | 20.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 70%|███████ | 20.7GB / 29.6GB, 199MB/s
|
||
New Data Upload : 99%|█████████▉| 20.3GB / 20.5GB, 199MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 70%|███████ | 20.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 70%|███████ | 20.7GB / 29.6GB, 197MB/s
|
||
New Data Upload : 99%|█████████▉| 20.3GB / 20.5GB, 197MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 70%|███████ | 20.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 70%|███████ | 20.8GB / 29.6GB, 196MB/s
|
||
New Data Upload : 99%|█████████▉| 20.3GB / 20.5GB, 196MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 70%|███████ | 20.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 70%|███████ | 20.8GB / 29.6GB, 196MB/s
|
||
New Data Upload : 99%|█████████▉| 20.4GB / 20.6GB, 196MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 71%|███████ | 20.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 70%|███████ | 20.8GB / 29.6GB, 195MB/s
|
||
New Data Upload : 99%|█████████▉| 20.4GB / 20.6GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 71%|███████ | 20.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 71%|███████ | 20.9GB / 29.6GB, 195MB/s
|
||
New Data Upload : 99%|█████████▉| 20.4GB / 20.7GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 71%|███████ | 20.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 71%|███████ | 20.9GB / 29.6GB, 195MB/s
|
||
New Data Upload : 99%|█████████▉| 20.5GB / 20.7GB, 195MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 71%|███████ | 20.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 71%|███████ | 20.9GB / 29.6GB, 192MB/s
|
||
New Data Upload : 99%|█████████▉| 20.5GB / 20.7GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 71%|███████ | 21.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 71%|███████ | 21.0GB / 29.6GB, 193MB/s
|
||
New Data Upload : 99%|█████████▉| 20.5GB / 20.7GB, 193MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 71%|███████ | 21.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 71%|███████ | 21.0GB / 29.6GB, 192MB/s
|
||
New Data Upload : 99%|█████████▉| 20.6GB / 20.8GB, 192MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 71%|███████ | 21.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 71%|███████ | 21.0GB / 29.6GB, 190MB/s
|
||
New Data Upload : 99%|█████████▉| 20.6GB / 20.8GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 71%|███████ | 21.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 71%|███████ | 21.1GB / 29.6GB, 188MB/s
|
||
New Data Upload : 99%|█████████▉| 20.6GB / 20.9GB, 188MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 71%|███████▏ | 21.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 71%|███████▏ | 21.1GB / 29.6GB, 185MB/s
|
||
New Data Upload : 99%|█████████▉| 20.6GB / 20.9GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 71%|███████▏ | 21.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 71%|███████▏ | 21.1GB / 29.6GB, 184MB/s
|
||
New Data Upload : 99%|█████████▉| 20.7GB / 20.9GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 72%|███████▏ | 21.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 72%|███████▏ | 21.1GB / 29.6GB, 184MB/s
|
||
New Data Upload : 99%|█████████▉| 20.7GB / 20.9GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 72%|███████▏ | 21.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 72%|███████▏ | 21.2GB / 29.6GB, 184MB/s
|
||
New Data Upload : 99%|█████████▉| 20.8GB / 21.0GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 72%|███████▏ | 21.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 72%|███████▏ | 21.2GB / 29.6GB, 186MB/s
|
||
New Data Upload : 99%|█████████▉| 20.8GB / 21.0GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 72%|███████▏ | 21.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 72%|███████▏ | 21.3GB / 29.6GB, 186MB/s
|
||
New Data Upload : 99%|█████████▉| 20.9GB / 21.1GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 72%|███████▏ | 21.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 72%|███████▏ | 21.3GB / 29.6GB, 186MB/s
|
||
New Data Upload : 99%|█████████▉| 20.9GB / 21.1GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 72%|███████▏ | 21.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 72%|███████▏ | 21.4GB / 29.6GB, 186MB/s
|
||
New Data Upload : 99%|█████████▉| 20.9GB / 21.1GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 72%|███████▏ | 21.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 72%|███████▏ | 21.4GB / 29.6GB, 187MB/s
|
||
New Data Upload : 99%|█████████▉| 21.0GB / 21.2GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 73%|███████▎ | 21.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 73%|███████▎ | 21.5GB / 29.6GB, 187MB/s
|
||
New Data Upload : 99%|█████████▉| 21.0GB / 21.2GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 73%|███████▎ | 21.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 73%|███████▎ | 21.5GB / 29.6GB, 186MB/s
|
||
New Data Upload : 99%|█████████▉| 21.1GB / 21.3GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 73%|███████▎ | 21.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 73%|███████▎ | 21.5GB / 29.6GB, 184MB/s
|
||
New Data Upload : 99%|█████████▉| 21.1GB / 21.3GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 73%|███████▎ | 21.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 73%|███████▎ | 21.5GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 21.1GB / 21.3GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 73%|███████▎ | 21.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 73%|███████▎ | 21.6GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 21.1GB / 21.3GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 73%|███████▎ | 21.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 73%|███████▎ | 21.6GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 21.2GB / 21.4GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 73%|███████▎ | 21.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 73%|███████▎ | 21.6GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 21.2GB / 21.5GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 73%|███████▎ | 21.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 73%|███████▎ | 21.7GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 21.2GB / 21.5GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 73%|███████▎ | 21.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 73%|███████▎ | 21.7GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 21.3GB / 21.5GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 74%|███████▎ | 21.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 74%|███████▎ | 21.7GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 21.3GB / 21.5GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 74%|███████▎ | 21.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 74%|███████▎ | 21.8GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▉| 21.3GB / 21.6GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 74%|███████▍ | 21.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 74%|███████▎ | 21.8GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▉| 21.4GB / 21.6GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 74%|███████▍ | 21.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 74%|███████▍ | 21.9GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 21.4GB / 21.7GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 74%|███████▍ | 21.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 74%|███████▍ | 21.9GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▉| 21.5GB / 21.7GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 74%|███████▍ | 21.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 74%|███████▍ | 21.9GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▉| 21.5GB / 21.7GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 74%|███████▍ | 22.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 74%|███████▍ | 22.0GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 21.5GB / 21.7GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 74%|███████▍ | 22.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 74%|███████▍ | 22.0GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 21.6GB / 21.8GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 75%|███████▍ | 22.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 75%|███████▍ | 22.1GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 21.6GB / 21.8GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 75%|███████▍ | 22.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 75%|███████▍ | 22.1GB / 29.6GB, 185MB/s
|
||
New Data Upload : 99%|█████████▉| 21.7GB / 21.9GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 75%|███████▍ | 22.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 75%|███████▍ | 22.1GB / 29.6GB, 185MB/s
|
||
New Data Upload : 99%|█████████▉| 21.7GB / 21.9GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 75%|███████▌ | 22.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 75%|███████▍ | 22.2GB / 29.6GB, 184MB/s
|
||
New Data Upload : 99%|█████████▉| 21.7GB / 21.9GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 75%|███████▌ | 22.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 75%|███████▌ | 22.2GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▉| 21.8GB / 21.9GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 75%|███████▌ | 22.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 75%|███████▌ | 22.2GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 21.8GB / 22.0GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 75%|███████▌ | 22.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 75%|███████▌ | 22.2GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 21.8GB / 22.0GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 75%|███████▌ | 22.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 75%|███████▌ | 22.3GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 21.8GB / 22.1GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 75%|███████▌ | 22.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 75%|███████▌ | 22.3GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 21.9GB / 22.1GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 76%|███████▌ | 22.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 76%|███████▌ | 22.3GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 21.9GB / 22.1GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 76%|███████▌ | 22.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 76%|███████▌ | 22.4GB / 29.6GB, 175MB/s
|
||
New Data Upload : 99%|█████████▉| 22.0GB / 22.1GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 76%|███████▌ | 22.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 76%|███████▌ | 22.4GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 22.0GB / 22.2GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 76%|███████▌ | 22.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 76%|███████▌ | 22.5GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 22.1GB / 22.2GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 76%|███████▌ | 22.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 76%|███████▌ | 22.5GB / 29.6GB, 176MB/s
|
||
New Data Upload : 99%|█████████▉| 22.1GB / 22.3GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 76%|███████▌ | 22.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 76%|███████▌ | 22.5GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 22.1GB / 22.3GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 76%|███████▋ | 22.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 76%|███████▌ | 22.5GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 22.1GB / 22.3GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 76%|███████▋ | 22.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 76%|███████▋ | 22.6GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 22.1GB / 22.3GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 76%|███████▋ | 22.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 76%|███████▋ | 22.6GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 22.2GB / 22.4GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 77%|███████▋ | 22.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 77%|███████▋ | 22.6GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 22.2GB / 22.4GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 77%|███████▋ | 22.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 77%|███████▋ | 22.7GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 22.2GB / 22.5GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 77%|███████▋ | 22.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 77%|███████▋ | 22.7GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 22.3GB / 22.5GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 77%|███████▋ | 22.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 77%|███████▋ | 22.7GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 22.3GB / 22.5GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 77%|███████▋ | 22.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 77%|███████▋ | 22.8GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 22.3GB / 22.6GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 77%|███████▋ | 22.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 77%|███████▋ | 22.8GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 22.4GB / 22.6GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 77%|███████▋ | 22.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 77%|███████▋ | 22.8GB / 29.6GB, 175MB/s
|
||
New Data Upload : 99%|█████████▉| 22.4GB / 22.7GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 77%|███████▋ | 22.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 77%|███████▋ | 22.9GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 22.5GB / 22.7GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 78%|███████▊ | 22.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 78%|███████▊ | 22.9GB / 29.6GB, 180MB/s
|
||
New Data Upload : 99%|█████████▉| 22.5GB / 22.7GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 78%|███████▊ | 23.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 78%|███████▊ | 23.0GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▉| 22.5GB / 22.7GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 78%|███████▊ | 23.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 78%|███████▊ | 23.0GB / 29.6GB, 180MB/s
|
||
New Data Upload : 99%|█████████▉| 22.6GB / 22.8GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 78%|███████▊ | 23.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 78%|███████▊ | 23.1GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▉| 22.6GB / 22.8GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 78%|███████▊ | 23.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 78%|███████▊ | 23.1GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▉| 22.7GB / 22.9GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 78%|███████▊ | 23.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 78%|███████▊ | 23.2GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▉| 22.7GB / 22.9GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 78%|███████▊ | 23.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 78%|███████▊ | 23.2GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 22.8GB / 22.9GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 79%|███████▊ | 23.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 79%|███████▊ | 23.2GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 22.8GB / 23.0GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 79%|███████▊ | 23.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 79%|███████▊ | 23.3GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 22.8GB / 23.1GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 79%|███████▉ | 23.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 79%|███████▉ | 23.3GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 22.9GB / 23.1GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 79%|███████▉ | 23.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 79%|███████▉ | 23.3GB / 29.6GB, 180MB/s
|
||
New Data Upload : 99%|█████████▉| 22.9GB / 23.1GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 79%|███████▉ | 23.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 79%|███████▉ | 23.4GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 22.9GB / 23.1GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 79%|███████▉ | 23.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 79%|███████▉ | 23.4GB / 29.6GB, 184MB/s
|
||
New Data Upload : 99%|█████████▉| 23.0GB / 23.1GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 79%|███████▉ | 23.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 79%|███████▉ | 23.5GB / 29.6GB, 184MB/s
|
||
New Data Upload : 99%|█████████▉| 23.0GB / 23.2GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 80%|███████▉ | 23.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 79%|███████▉ | 23.5GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 23.1GB / 23.3GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 80%|███████▉ | 23.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 80%|███████▉ | 23.5GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▉| 23.1GB / 23.3GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 80%|███████▉ | 23.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 80%|███████▉ | 23.6GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 23.1GB / 23.3GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 80%|███████▉ | 23.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 80%|███████▉ | 23.6GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 23.2GB / 23.3GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 80%|███████▉ | 23.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 80%|███████▉ | 23.6GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 23.2GB / 23.4GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 80%|████████ | 23.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 80%|████████ | 23.7GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▉| 23.2GB / 23.5GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 80%|████████ | 23.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 80%|████████ | 23.7GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▉| 23.2GB / 23.5GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 80%|████████ | 23.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 80%|████████ | 23.7GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▉| 23.3GB / 23.5GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 80%|████████ | 23.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 80%|████████ | 23.7GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 23.3GB / 23.5GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 80%|████████ | 23.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 80%|████████ | 23.8GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 23.3GB / 23.5GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 81%|████████ | 23.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 81%|████████ | 23.8GB / 29.6GB, 176MB/s
|
||
New Data Upload : 99%|█████████▉| 23.4GB / 23.5GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 81%|████████ | 23.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 81%|████████ | 23.8GB / 29.6GB, 175MB/s
|
||
New Data Upload : 99%|█████████▉| 23.4GB / 23.6GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 81%|████████ | 23.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 81%|████████ | 23.9GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 23.4GB / 23.6GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 81%|████████ | 23.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 81%|████████ | 23.9GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 23.5GB / 23.6GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 81%|████████ | 23.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 81%|████████ | 23.9GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 23.5GB / 23.7GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 81%|████████ | 24.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 81%|████████ | 24.0GB / 29.6GB, 175MB/s
|
||
New Data Upload : 99%|█████████▉| 23.5GB / 23.7GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 81%|████████ | 24.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 81%|████████ | 24.0GB / 29.6GB, 176MB/s
|
||
New Data Upload : 99%|█████████▉| 23.6GB / 23.7GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 81%|████████▏ | 24.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 81%|████████▏ | 24.0GB / 29.6GB, 176MB/s
|
||
New Data Upload : 99%|█████████▉| 23.6GB / 23.8GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 81%|████████▏ | 24.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 81%|████████▏ | 24.1GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 23.6GB / 23.8GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 82%|████████▏ | 24.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 82%|████████▏ | 24.1GB / 29.6GB, 176MB/s
|
||
New Data Upload : 99%|█████████▉| 23.7GB / 23.9GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 82%|████████▏ | 24.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 82%|████████▏ | 24.1GB / 29.6GB, 175MB/s
|
||
New Data Upload : 99%|█████████▉| 23.7GB / 23.9GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 82%|████████▏ | 24.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 82%|████████▏ | 24.2GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 23.7GB / 23.9GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 82%|████████▏ | 24.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 82%|████████▏ | 24.2GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 23.8GB / 24.0GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 82%|████████▏ | 24.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 82%|████████▏ | 24.2GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 23.8GB / 24.0GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 82%|████████▏ | 24.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 82%|████████▏ | 24.3GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 23.8GB / 24.1GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 82%|████████▏ | 24.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 82%|████████▏ | 24.3GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 23.9GB / 24.1GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 82%|████████▏ | 24.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 82%|████████▏ | 24.3GB / 29.6GB, 176MB/s
|
||
New Data Upload : 99%|█████████▉| 23.9GB / 24.1GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 82%|████████▏ | 24.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 82%|████████▏ | 24.4GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 23.9GB / 24.1GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 83%|████████▎ | 24.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 83%|████████▎ | 24.4GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 24.0GB / 24.2GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 83%|████████▎ | 24.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 83%|████████▎ | 24.4GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 24.0GB / 24.2GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 83%|████████▎ | 24.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 83%|████████▎ | 24.5GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 24.0GB / 24.3GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 83%|████████▎ | 24.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 83%|████████▎ | 24.5GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 24.1GB / 24.3GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 83%|████████▎ | 24.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 83%|████████▎ | 24.6GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▉| 24.1GB / 24.3GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 83%|████████▎ | 24.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 83%|████████▎ | 24.6GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 24.2GB / 24.3GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 83%|████████▎ | 24.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 83%|████████▎ | 24.6GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 24.2GB / 24.4GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 83%|████████▎ | 24.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 83%|████████▎ | 24.6GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 24.2GB / 24.4GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 83%|████████▎ | 24.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 83%|████████▎ | 24.7GB / 29.6GB, 175MB/s
|
||
New Data Upload : 99%|█████████▉| 24.2GB / 24.5GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 84%|████████▎ | 24.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 84%|████████▎ | 24.7GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 24.3GB / 24.5GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 84%|████████▎ | 24.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 84%|████████▎ | 24.7GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 24.3GB / 24.5GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 84%|████████▍ | 24.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 84%|████████▍ | 24.8GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 24.4GB / 24.5GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 84%|████████▍ | 24.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 84%|████████▍ | 24.8GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 24.4GB / 24.6GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 84%|████████▍ | 24.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 84%|████████▍ | 24.9GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 24.5GB / 24.7GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 84%|████████▍ | 24.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 84%|████████▍ | 24.9GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 24.5GB / 24.7GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 84%|████████▍ | 25.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 84%|████████▍ | 25.0GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 24.5GB / 24.7GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 85%|████████▍ | 25.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 85%|████████▍ | 25.0GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 24.6GB / 24.7GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 85%|████████▍ | 25.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 85%|████████▍ | 25.0GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 24.6GB / 24.8GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 85%|████████▍ | 25.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 85%|████████▍ | 25.1GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 24.6GB / 24.9GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 85%|████████▍ | 25.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 85%|████████▍ | 25.1GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 24.7GB / 24.9GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 85%|████████▌ | 25.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 85%|████████▍ | 25.1GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 24.7GB / 24.9GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 85%|████████▌ | 25.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 85%|████████▌ | 25.2GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 24.7GB / 24.9GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 85%|████████▌ | 25.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 85%|████████▌ | 25.2GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 24.8GB / 25.0GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 85%|████████▌ | 25.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 85%|████████▌ | 25.3GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 24.8GB / 25.1GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 86%|████████▌ | 25.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 86%|████████▌ | 25.3GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 24.9GB / 25.1GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 86%|████████▌ | 25.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 86%|████████▌ | 25.3GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 24.9GB / 25.1GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 86%|████████▌ | 25.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 86%|████████▌ | 25.4GB / 29.6GB, 175MB/s
|
||
New Data Upload : 99%|█████████▉| 24.9GB / 25.2GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 86%|████████▌ | 25.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 86%|████████▌ | 25.4GB / 29.6GB, 175MB/s
|
||
New Data Upload : 99%|█████████▉| 25.0GB / 25.2GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 86%|████████▌ | 25.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 86%|████████▌ | 25.5GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▉| 25.0GB / 25.3GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 86%|████████▋ | 25.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 86%|████████▋ | 25.5GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▉| 25.1GB / 25.3GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 87%|████████▋ | 25.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 86%|████████▋ | 25.6GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 25.1GB / 25.3GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 87%|████████▋ | 25.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 87%|████████▋ | 25.6GB / 29.6GB, 184MB/s
|
||
New Data Upload : 99%|█████████▉| 25.2GB / 25.3GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 87%|████████▋ | 25.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 87%|████████▋ | 25.7GB / 29.6GB, 185MB/s
|
||
New Data Upload : 99%|█████████▉| 25.2GB / 25.4GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 87%|████████▋ | 25.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 87%|████████▋ | 25.7GB / 29.6GB, 184MB/s
|
||
New Data Upload : 99%|█████████▉| 25.3GB / 25.4GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 87%|████████▋ | 25.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 87%|████████▋ | 25.7GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 25.3GB / 25.5GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 87%|████████▋ | 25.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 87%|████████▋ | 25.7GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 25.3GB / 25.5GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 87%|████████▋ | 25.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 87%|████████▋ | 25.8GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 25.3GB / 25.6GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 87%|████████▋ | 25.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 87%|████████▋ | 25.8GB / 29.6GB, 184MB/s
|
||
New Data Upload : 99%|█████████▉| 25.4GB / 25.6GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 87%|████████▋ | 25.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 87%|████████▋ | 25.8GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 25.4GB / 25.6GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 88%|████████▊ | 25.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 88%|████████▊ | 25.9GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 25.4GB / 25.7GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 88%|████████▊ | 25.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 88%|████████▊ | 25.9GB / 29.6GB, 184MB/s
|
||
New Data Upload : 99%|█████████▉| 25.5GB / 25.7GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 88%|████████▊ | 25.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 88%|████████▊ | 25.9GB / 29.6GB, 184MB/s
|
||
New Data Upload : 99%|█████████▉| 25.5GB / 25.8GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 88%|████████▊ | 26.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 88%|████████▊ | 26.0GB / 29.6GB, 184MB/s
|
||
New Data Upload : 99%|█████████▉| 25.5GB / 25.8GB, 184MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 88%|████████▊ | 26.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 88%|████████▊ | 26.0GB / 29.6GB, 185MB/s
|
||
New Data Upload : 99%|█████████▉| 25.6GB / 25.8GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 88%|████████▊ | 26.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 88%|████████▊ | 26.1GB / 29.6GB, 185MB/s
|
||
New Data Upload : 99%|█████████▉| 25.6GB / 25.9GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 88%|████████▊ | 26.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 88%|████████▊ | 26.1GB / 29.6GB, 185MB/s
|
||
New Data Upload : 99%|█████████▉| 25.7GB / 25.9GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 88%|████████▊ | 26.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 88%|████████▊ | 26.1GB / 29.6GB, 187MB/s
|
||
New Data Upload : 99%|█████████▉| 25.7GB / 26.0GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 89%|████████▊ | 26.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 89%|████████▊ | 26.2GB / 29.6GB, 188MB/s
|
||
New Data Upload : 99%|█████████▉| 25.8GB / 26.0GB, 188MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 89%|████████▉ | 26.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 89%|████████▊ | 26.2GB / 29.6GB, 190MB/s
|
||
New Data Upload : 99%|█████████▉| 25.8GB / 26.0GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 89%|████████▉ | 26.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 89%|████████▉ | 26.3GB / 29.6GB, 191MB/s
|
||
New Data Upload : 99%|█████████▉| 25.9GB / 26.0GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 89%|████████▉ | 26.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 89%|████████▉ | 26.3GB / 29.6GB, 191MB/s
|
||
New Data Upload : 99%|█████████▉| 25.9GB / 26.1GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 89%|████████▉ | 26.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 89%|████████▉ | 26.4GB / 29.6GB, 191MB/s
|
||
New Data Upload : 99%|█████████▉| 25.9GB / 26.1GB, 191MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 89%|████████▉ | 26.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 89%|████████▉ | 26.4GB / 29.6GB, 190MB/s
|
||
New Data Upload : 99%|█████████▉| 25.9GB / 26.2GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 89%|████████▉ | 26.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 89%|████████▉ | 26.4GB / 29.6GB, 189MB/s
|
||
New Data Upload : 99%|█████████▉| 26.0GB / 26.2GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 89%|████████▉ | 26.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 89%|████████▉ | 26.4GB / 29.6GB, 188MB/s
|
||
New Data Upload : 99%|█████████▉| 26.0GB / 26.2GB, 188MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 89%|████████▉ | 26.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 89%|████████▉ | 26.4GB / 29.6GB, 186MB/s
|
||
New Data Upload : 99%|█████████▉| 26.0GB / 26.2GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 90%|████████▉ | 26.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 90%|████████▉ | 26.5GB / 29.6GB, 186MB/s
|
||
New Data Upload : 99%|█████████▉| 26.0GB / 26.2GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 90%|████████▉ | 26.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 90%|████████▉ | 26.5GB / 29.6GB, 186MB/s
|
||
New Data Upload : 99%|█████████▉| 26.1GB / 26.3GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 90%|████████▉ | 26.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 90%|████████▉ | 26.5GB / 29.6GB, 186MB/s
|
||
New Data Upload : 99%|█████████▉| 26.1GB / 26.3GB, 186MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 90%|████████▉ | 26.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 90%|████████▉ | 26.6GB / 29.6GB, 187MB/s
|
||
New Data Upload : 99%|█████████▉| 26.2GB / 26.4GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 90%|█████████ | 26.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 90%|█████████ | 26.6GB / 29.6GB, 190MB/s
|
||
New Data Upload : 99%|█████████▉| 26.2GB / 26.4GB, 190MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 90%|█████████ | 26.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 90%|█████████ | 26.7GB / 29.6GB, 189MB/s
|
||
New Data Upload : 99%|█████████▉| 26.2GB / 26.4GB, 189MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 90%|█████████ | 26.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 90%|█████████ | 26.7GB / 29.6GB, 188MB/s
|
||
New Data Upload : 99%|█████████▉| 26.3GB / 26.5GB, 188MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 90%|█████████ | 26.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 90%|█████████ | 26.7GB / 29.6GB, 187MB/s
|
||
New Data Upload : 99%|█████████▉| 26.3GB / 26.5GB, 187MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 91%|█████████ | 26.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 91%|█████████ | 26.8GB / 29.6GB, 185MB/s
|
||
New Data Upload : 99%|█████████▉| 26.3GB / 26.5GB, 185MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 91%|█████████ | 26.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 91%|█████████ | 26.8GB / 29.6GB, 183MB/s
|
||
New Data Upload : 99%|█████████▉| 26.4GB / 26.6GB, 183MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 91%|█████████ | 26.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 91%|█████████ | 26.8GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▉| 26.4GB / 26.6GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 91%|█████████ | 26.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 91%|█████████ | 26.8GB / 29.6GB, 180MB/s
|
||
New Data Upload : 99%|█████████▉| 26.4GB / 26.6GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 91%|█████████ | 26.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 91%|█████████ | 26.9GB / 29.6GB, 180MB/s
|
||
New Data Upload : 99%|█████████▉| 26.4GB / 26.7GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 91%|█████████ | 26.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 91%|█████████ | 26.9GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▉| 26.5GB / 26.7GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 91%|█████████ | 26.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 91%|█████████ | 26.9GB / 29.6GB, 180MB/s
|
||
New Data Upload : 99%|█████████▉| 26.5GB / 26.8GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 91%|█████████ | 26.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 91%|█████████ | 27.0GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▉| 26.5GB / 26.8GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 91%|█████████▏| 27.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 91%|█████████▏| 27.0GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 26.6GB / 26.8GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 91%|█████████▏| 27.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 91%|█████████▏| 27.0GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 26.6GB / 26.8GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 92%|█████████▏| 27.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 92%|█████████▏| 27.1GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 26.6GB / 26.9GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 92%|█████████▏| 27.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 92%|█████████▏| 27.1GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 26.7GB / 26.9GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 92%|█████████▏| 27.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 92%|█████████▏| 27.1GB / 29.6GB, 176MB/s
|
||
New Data Upload : 99%|█████████▉| 26.7GB / 26.9GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 92%|█████████▏| 27.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 92%|█████████▏| 27.2GB / 29.6GB, 175MB/s
|
||
New Data Upload : 99%|█████████▉| 26.7GB / 27.0GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 92%|█████████▏| 27.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 92%|█████████▏| 27.2GB / 29.6GB, 175MB/s
|
||
New Data Upload : 99%|█████████▉| 26.8GB / 27.0GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 92%|█████████▏| 27.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 92%|█████████▏| 27.3GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 26.8GB / 27.0GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 92%|█████████▏| 27.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 92%|█████████▏| 27.3GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 26.9GB / 27.0GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 92%|█████████▏| 27.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 92%|█████████▏| 27.3GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 26.9GB / 27.1GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 93%|█████████▎| 27.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 92%|█████████▏| 27.3GB / 29.6GB, 169MB/s
|
||
New Data Upload : 99%|█████████▉| 26.9GB / 27.1GB, 169MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 93%|█████████▎| 27.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 93%|█████████▎| 27.4GB / 29.6GB, 168MB/s
|
||
New Data Upload : 99%|█████████▉| 26.9GB / 27.2GB, 168MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 93%|█████████▎| 27.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 93%|█████████▎| 27.4GB / 29.6GB, 169MB/s
|
||
New Data Upload : 99%|█████████▉| 27.0GB / 27.2GB, 169MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 93%|█████████▎| 27.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 93%|█████████▎| 27.5GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 27.0GB / 27.2GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 93%|█████████▎| 27.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 93%|█████████▎| 27.5GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 27.1GB / 27.3GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 93%|█████████▎| 27.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 93%|█████████▎| 27.5GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 27.1GB / 27.3GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 93%|█████████▎| 27.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 93%|█████████▎| 27.6GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 27.1GB / 27.4GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 93%|█████████▎| 27.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 93%|█████████▎| 27.6GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 27.2GB / 27.4GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 94%|█████████▎| 27.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 93%|█████████▎| 27.6GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 27.2GB / 27.4GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 94%|█████████▎| 27.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 94%|█████████▎| 27.7GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 27.3GB / 27.5GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 94%|█████████▍| 27.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 94%|█████████▎| 27.7GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 27.3GB / 27.5GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 94%|█████████▍| 27.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 94%|█████████▍| 27.8GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 27.3GB / 27.6GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 94%|█████████▍| 27.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 94%|█████████▍| 27.8GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 27.3GB / 27.6GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 94%|█████████▍| 27.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 94%|█████████▍| 27.8GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 27.4GB / 27.6GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 94%|█████████▍| 27.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 94%|█████████▍| 27.9GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 27.4GB / 27.6GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 94%|█████████▍| 27.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 94%|█████████▍| 27.9GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 27.5GB / 27.7GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 95%|█████████▍| 27.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 94%|█████████▍| 27.9GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 27.5GB / 27.7GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 95%|█████████▍| 28.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 95%|█████████▍| 28.0GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 27.5GB / 27.8GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 95%|█████████▍| 28.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 95%|█████████▍| 28.0GB / 29.6GB, 169MB/s
|
||
New Data Upload : 99%|█████████▉| 27.6GB / 27.8GB, 169MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 95%|█████████▍| 28.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 95%|█████████▍| 28.1GB / 29.6GB, 170MB/s
|
||
New Data Upload : 99%|█████████▉| 27.6GB / 27.8GB, 170MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 95%|█████████▌| 28.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 95%|█████████▌| 28.1GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 27.7GB / 27.9GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 95%|█████████▌| 28.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 95%|█████████▌| 28.1GB / 29.6GB, 171MB/s
|
||
New Data Upload : 99%|█████████▉| 27.7GB / 28.0GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 95%|█████████▌| 28.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 95%|█████████▌| 28.2GB / 29.6GB, 172MB/s
|
||
New Data Upload : 99%|█████████▉| 27.7GB / 28.0GB, 172MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 95%|█████████▌| 28.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 95%|█████████▌| 28.2GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 27.8GB / 28.0GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 96%|█████████▌| 28.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 96%|█████████▌| 28.2GB / 29.6GB, 176MB/s
|
||
New Data Upload : 99%|█████████▉| 27.8GB / 28.0GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 96%|█████████▌| 28.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 96%|█████████▌| 28.3GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 27.8GB / 28.1GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 96%|█████████▌| 28.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 96%|█████████▌| 28.3GB / 29.6GB, 176MB/s
|
||
New Data Upload : 99%|█████████▉| 27.9GB / 28.2GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 96%|█████████▌| 28.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 96%|█████████▌| 28.4GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 27.9GB / 28.2GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 96%|█████████▌| 28.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 96%|█████████▌| 28.4GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 28.0GB / 28.2GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 96%|█████████▌| 28.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 96%|█████████▌| 28.4GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 28.0GB / 28.2GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 96%|█████████▋| 28.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 96%|█████████▋| 28.5GB / 29.6GB, 180MB/s
|
||
New Data Upload : 99%|█████████▉| 28.1GB / 28.3GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 97%|█████████▋| 28.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 97%|█████████▋| 28.5GB / 29.6GB, 180MB/s
|
||
New Data Upload : 99%|█████████▉| 28.1GB / 28.3GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 97%|█████████▋| 28.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 97%|█████████▋| 28.6GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▉| 28.1GB / 28.3GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 97%|█████████▋| 28.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 97%|█████████▋| 28.6GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 28.2GB / 28.4GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 97%|█████████▋| 28.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 97%|█████████▋| 28.6GB / 29.6GB, 180MB/s
|
||
New Data Upload : 99%|█████████▉| 28.2GB / 28.4GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 97%|█████████▋| 28.6GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 97%|█████████▋| 28.7GB / 29.6GB, 180MB/s
|
||
New Data Upload : 99%|█████████▉| 28.2GB / 28.4GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 97%|█████████▋| 28.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 97%|█████████▋| 28.7GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▉| 28.2GB / 28.4GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 97%|█████████▋| 28.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 97%|█████████▋| 28.7GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▉| 28.3GB / 28.4GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 97%|█████████▋| 28.7GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 97%|█████████▋| 28.7GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 28.3GB / 28.5GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 97%|█████████▋| 28.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 97%|█████████▋| 28.8GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 28.4GB / 28.5GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 98%|█████████▊| 28.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 97%|█████████▋| 28.8GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 28.4GB / 28.6GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 98%|█████████▊| 28.8GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 98%|█████████▊| 28.9GB / 29.6GB, 182MB/s
|
||
New Data Upload : 99%|█████████▉| 28.4GB / 28.6GB, 182MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 98%|█████████▊| 28.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 98%|█████████▊| 28.9GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▉| 28.4GB / 28.6GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 98%|█████████▊| 28.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 98%|█████████▊| 28.9GB / 29.6GB, 181MB/s
|
||
New Data Upload : 99%|█████████▉| 28.5GB / 28.6GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 98%|█████████▊| 28.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 98%|█████████▊| 28.9GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▉| 28.5GB / 28.7GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 98%|█████████▊| 28.9GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 98%|█████████▊| 29.0GB / 29.6GB, 179MB/s
|
||
New Data Upload : 99%|█████████▉| 28.5GB / 28.7GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 98%|█████████▊| 29.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 98%|█████████▊| 29.0GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 28.5GB / 28.8GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 98%|█████████▊| 29.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 98%|█████████▊| 29.0GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 28.6GB / 28.8GB, 177MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 98%|█████████▊| 29.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 98%|█████████▊| 29.0GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 28.6GB / 28.8GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 98%|█████████▊| 29.0GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 98%|█████████▊| 29.1GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 28.6GB / 28.9GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 98%|█████████▊| 29.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 98%|█████████▊| 29.1GB / 29.6GB, 173MB/s
|
||
New Data Upload : 99%|█████████▉| 28.6GB / 28.9GB, 173MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 99%|█████████▊| 29.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 98%|█████████▊| 29.1GB / 29.6GB, 174MB/s
|
||
New Data Upload : 99%|█████████▉| 28.7GB / 29.0GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 99%|█████████▊| 29.1GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 99%|█████████▊| 29.2GB / 29.6GB, 175MB/s
|
||
New Data Upload : 99%|█████████▉| 28.7GB / 29.0GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 99%|█████████▉| 29.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 99%|█████████▉| 29.2GB / 29.6GB, 176MB/s
|
||
New Data Upload : 99%|█████████▉| 28.8GB / 29.0GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 99%|█████████▉| 29.2GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 99%|█████████▉| 29.2GB / 29.6GB, 175MB/s
|
||
New Data Upload : 99%|█████████▉| 28.8GB / 29.1GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 99%|█████████▉| 29.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 99%|█████████▉| 29.3GB / 29.6GB, 175MB/s
|
||
New Data Upload : 99%|█████████▉| 28.9GB / 29.1GB, 175MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 99%|█████████▉| 29.3GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 33%|███▎ | 7.23MB / 22.2MB [A[A[A[A
Processing Files (1 / 3) : 99%|█████████▉| 29.3GB / 29.6GB, 177MB/s
|
||
New Data Upload : 99%|█████████▉| 28.9GB / 29.1GB, 177MB/s [A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 20%|██ | 1.50kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 99%|█████████▉| 29.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 46%|████▋ | 10.3MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 20%|██ | 1.50kB / 7.38kB [A[A[A[A[A
Processing Files (1 / 4) : 99%|█████████▉| 29.4GB / 29.6GB, 178MB/s
|
||
New Data Upload : 99%|█████████▉| 29.0GB / 29.1GB, 178MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.4GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 54%|█████▍ | 12.0MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 32%|███▏ | 2.36kB / 7.38kB [A[A[A[A[A
Processing Files (1 / 4) : 100%|█████████▉| 29.4GB / 29.6GB, 180MB/s
|
||
New Data Upload : 100%|█████████▉| 29.0GB / 29.1GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 70%|██████▉ | 15.5MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 55%|█████▌ | 4.08kB / 7.38kB [A[A[A[A[A
Processing Files (1 / 4) : 100%|█████████▉| 29.5GB / 29.6GB, 181MB/s
|
||
New Data Upload : 100%|█████████▉| 29.1GB / 29.1GB, 181MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 89%|████████▉ | 19.9MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 84%|████████▍ | 6.22kB / 7.38kB [A[A[A[A[A
Processing Files (1 / 4) : 100%|█████████▉| 29.5GB / 29.6GB, 180MB/s
|
||
New Data Upload : 100%|█████████▉| 29.1GB / 29.1GB, 180MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 99%|█████████▉| 22.1MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 99%|█████████▉| 7.30kB / 7.38kB [A[A[A[A[A
Processing Files (1 / 4) : 100%|█████████▉| 29.5GB / 29.6GB, 179MB/s
|
||
New Data Upload : 100%|█████████▉| 29.1GB / 29.1GB, 179MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 99%|█████████▉| 22.1MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 99%|█████████▉| 7.30kB / 7.38kB [A[A[A[A[A
Processing Files (1 / 4) : 100%|█████████▉| 29.5GB / 29.6GB, 176MB/s
|
||
New Data Upload : 100%|█████████▉| 29.1GB / 29.1GB, 176MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 99%|█████████▉| 22.1MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 99%|█████████▉| 7.30kB / 7.38kB [A[A[A[A[A
Processing Files (1 / 4) : 100%|█████████▉| 29.6GB / 29.6GB, 174MB/s
|
||
New Data Upload : 100%|█████████▉| 29.1GB / 29.1GB, 174MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
Processing Files (3 / 4) : 100%|█████████▉| 29.6GB / 29.6GB, 171MB/s
|
||
New Data Upload : 100%|█████████▉| 29.1GB / 29.1GB, 171MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
Processing Files (3 / 4) : 100%|█████████▉| 29.6GB / 29.6GB, 168MB/s
|
||
New Data Upload : 100%|█████████▉| 29.1GB / 29.1GB, 168MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
Processing Files (4 / 4) : 100%|██████████| 29.6GB / 29.6GB, 153MB/s
|
||
New Data Upload : 100%|██████████| 29.1GB / 29.1GB, 153MB/s [A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A[A[A[A
Processing Files (4 / 4) : 100%|██████████| 29.6GB / 29.6GB, 97.5MB/s
|
||
New Data Upload : 100%|██████████| 29.1GB / 29.1GB, 97.5MB/s
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB
|
||
***** train metrics *****
|
||
epoch = 4.0
|
||
total_flos = 141876GF
|
||
train_loss = 0.4443
|
||
train_runtime = 5:16:28.05
|
||
train_samples_per_second = 6.32
|
||
train_steps_per_second = 0.05
|
||
Figure saved at: saves/qwen3_14b/coding/fft/training_loss.png
|
||
[WARNING|2026-04-30 21:42:59] llamafactory.extras.ploting:149 >> No metric eval_loss to plot.
|
||
[WARNING|2026-04-30 21:42:59] llamafactory.extras.ploting:149 >> No metric eval_accuracy to plot.
|
||
[INFO|trainer.py:3797] 2026-04-30 21:43:18,948 >> Saving model checkpoint to saves/qwen3_14b/coding/fft
|
||
[INFO|configuration_utils.py:432] 2026-04-30 21:43:18,956 >> Configuration saved in saves/qwen3_14b/coding/fft/config.json
|
||
[INFO|configuration_utils.py:803] 2026-04-30 21:43:18,960 >> Configuration saved in saves/qwen3_14b/coding/fft/generation_config.json
|
||
Writing model shards: 0%| | 0/1 [00:00<?, ?it/s]ywang29-p4d-debug-2-worker-0:83701:88967 [5] NCCL INFO comm 0x55dd0ac0a840 rank 5 nranks 8 cudaDev 5 busId 901d0 - Destroy COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83702:88963 [6] NCCL INFO comm 0x55a9dbb4c8b0 rank 6 nranks 8 cudaDev 6 busId a01c0 - Destroy COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83698:88971 [2] NCCL INFO comm 0x55e5220b9800 rank 2 nranks 8 cudaDev 2 busId 201c0 - Destroy COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83699:88972 [3] NCCL INFO comm 0x5608d6d641f0 rank 3 nranks 8 cudaDev 3 busId 201d0 - Destroy COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83700:88964 [4] NCCL INFO comm 0x562ca54bbc00 rank 4 nranks 8 cudaDev 4 busId 901c0 - Destroy COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83701:88974 [5] NCCL INFO comm 0x55dcff54b080 rank 5 nranks 8 cudaDev 5 busId 901d0 - Destroy COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83699:88980 [3] NCCL INFO comm 0x5608cb6aab30 rank 3 nranks 8 cudaDev 3 busId 201d0 - Destroy COMPLETE
|
||
ywang29-p4d-debug-2-worker-0:83700:88982 [4] NCCL INFO comm 0x562c99db2980 rank 4 nranks 8 cudaDev 4 busId 901c0 - Destroy COMPLETE
|
||
[W430 21:43:21.952781237 AllocatorConfig.cpp:28] Warning: PYTORCH_CUDA_ALLOC_CONF is deprecated, use PYTORCH_ALLOC_CONF instead (function operator())
|
||
[W430 21:43:21.234194573 AllocatorConfig.cpp:28] Warning: PYTORCH_CUDA_ALLOC_CONF is deprecated, use PYTORCH_ALLOC_CONF instead (function operator())
|
||
[W430 21:43:21.305356990 AllocatorConfig.cpp:28] Warning: PYTORCH_CUDA_ALLOC_CONF is deprecated, use PYTORCH_ALLOC_CONF instead (function operator())
|
||
Writing model shards: 100%|██████████| 1/1 [00:50<00:00, 50.90s/it]
Writing model shards: 100%|██████████| 1/1 [00:50<00:00, 50.90s/it]
|
||
[INFO|modeling_utils.py:3380] 2026-04-30 21:44:09,906 >> Model weights saved in saves/qwen3_14b/coding/fft/model.safetensors
|
||
[INFO|tokenization_utils_base.py:3224] 2026-04-30 21:44:09,912 >> chat template saved in saves/qwen3_14b/coding/fft/chat_template.jinja
|
||
[INFO|tokenization_utils_base.py:2078] 2026-04-30 21:44:09,916 >> tokenizer config file saved in saves/qwen3_14b/coding/fft/tokenizer_config.json
|
||
[INFO|modelcard.py:266] 2026-04-30 21:44:11,147 >> Dropping the following result as it does not have all the necessary fields:
|
||
{'task': {'name': 'Causal Language Modeling', 'type': 'text-generation'}}
|
||
Processing Files (0 / 0) : | | 0.00B / 0.00B
|
||
New Data Upload : | | 0.00B / 0.00B [A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 15.9MB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 0%| | 15.9MB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 0%| | 49.6MB / 29.6GB, ???B/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 1%| | 184MB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 1%| | 218MB / 29.6GB, 840MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 1%| | 352MB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 1%|▏ | 386MB / 29.6GB, 840MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 2%|▏ | 512MB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 2%|▏ | 546MB / 29.6GB, 826MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 2%|▏ | 688MB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 2%|▏ | 722MB / 29.6GB, 840MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 3%|▎ | 864MB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 3%|▎ | 898MB / 29.6GB, 848MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 4%|▎ | 1.04GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 4%|▎ | 1.07GB / 29.6GB, 853MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 4%|▍ | 1.22GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 4%|▍ | 1.25GB / 29.6GB, 857MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 5%|▍ | 1.38GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 5%|▍ | 1.41GB / 29.6GB, 850MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 5%|▌ | 1.54GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 5%|▌ | 1.58GB / 29.6GB, 849MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 6%|▌ | 1.71GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 6%|▌ | 1.75GB / 29.6GB, 848MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 6%|▋ | 1.89GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 6%|▋ | 1.92GB / 29.6GB, 851MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 7%|▋ | 2.06GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 7%|▋ | 2.10GB / 29.6GB, 853MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 8%|▊ | 2.25GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 8%|▊ | 2.28GB / 29.6GB, 858MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 8%|▊ | 2.42GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 8%|▊ | 2.46GB / 29.6GB, 860MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 9%|▉ | 2.59GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 9%|▉ | 2.63GB / 29.6GB, 859MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 9%|▉ | 2.75GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 9%|▉ | 2.79GB / 29.6GB, 855MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 10%|▉ | 2.92GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 10%|▉ | 2.96GB / 29.6GB, 855MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 10%|█ | 3.08GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 11%|█ | 3.11GB / 29.6GB, 850MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 11%|█ | 3.23GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 11%|█ | 3.27GB / 29.6GB, 847MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 11%|█▏ | 3.39GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 12%|█▏ | 3.43GB / 29.6GB, 845MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 12%|█▏ | 3.56GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 12%|█▏ | 3.60GB / 29.6GB, 844MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 13%|█▎ | 3.72GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 13%|█▎ | 3.76GB / 29.6GB, 842MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 13%|█▎ | 3.88GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 13%|█▎ | 3.92GB / 29.6GB, 841MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 14%|█▎ | 4.06GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 14%|█▍ | 4.09GB / 29.6GB, 842MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 14%|█▍ | 4.23GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 14%|█▍ | 4.26GB / 29.6GB, 842MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 15%|█▍ | 4.34GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 15%|█▍ | 4.37GB / 29.6GB, 831MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 15%|█▌ | 4.51GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 15%|█▌ | 4.54GB / 29.6GB, 832MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 16%|█▌ | 4.67GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 16%|█▌ | 4.71GB / 29.6GB, 832MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 16%|█▋ | 4.84GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 16%|█▋ | 4.88GB / 29.6GB, 832MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 17%|█▋ | 5.02GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 17%|█▋ | 5.05GB / 29.6GB, 834MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 18%|█▊ | 5.19GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 18%|█▊ | 5.23GB / 29.6GB, 835MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 18%|█▊ | 5.36GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 18%|█▊ | 5.40GB / 29.6GB, 835MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 19%|█▉ | 5.55GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 19%|█▉ | 5.58GB / 29.6GB, 838MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 19%|█▉ | 5.73GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 19%|█▉ | 5.76GB / 29.6GB, 840MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 20%|██ | 5.91GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 20%|██ | 5.95GB / 29.6GB, 843MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 21%|██ | 6.10GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 21%|██ | 6.13GB / 29.6GB, 845MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 21%|██ | 6.27GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 21%|██▏ | 6.30GB / 29.6GB, 845MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 22%|██▏ | 6.45GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 22%|██▏ | 6.48GB / 29.6GB, 847MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 22%|██▏ | 6.63GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 23%|██▎ | 6.67GB / 29.6GB, 849MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 23%|██▎ | 6.82GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 23%|██▎ | 6.85GB / 29.6GB, 850MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 24%|██▎ | 7.00GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 24%|██▍ | 7.04GB / 29.6GB, 852MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 24%|██▍ | 7.19GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 24%|██▍ | 7.22GB / 29.6GB, 854MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 25%|██▍ | 7.35GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 25%|██▍ | 7.39GB / 29.6GB, 853MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 26%|██▌ | 7.54GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 26%|██▌ | 7.57GB / 29.6GB, 855MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 26%|██▌ | 7.72GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 26%|██▌ | 7.76GB / 29.6GB, 856MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 27%|██▋ | 7.91GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 27%|██▋ | 7.94GB / 29.6GB, 858MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 27%|██▋ | 8.09GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 27%|██▋ | 8.12GB / 29.6GB, 859MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 28%|██▊ | 8.27GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 28%|██▊ | 8.31GB / 29.6GB, 860MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 29%|██▊ | 8.46GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 29%|██▊ | 8.49GB / 29.6GB, 861MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 29%|██▉ | 8.64GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 29%|██▉ | 8.68GB / 29.6GB, 863MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 30%|██▉ | 8.83GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 30%|██▉ | 8.86GB / 29.6GB, 864MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 30%|███ | 9.00GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 31%|███ | 9.04GB / 29.6GB, 865MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 31%|███ | 9.19GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 31%|███ | 9.22GB / 29.6GB, 866MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 32%|███▏ | 9.37GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 32%|███▏ | 9.40GB / 29.6GB, 869MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 32%|███▏ | 9.55GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 32%|███▏ | 9.59GB / 29.6GB, 869MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 33%|███▎ | 9.74GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 33%|███▎ | 9.77GB / 29.6GB, 870MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 34%|███▎ | 9.93GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 34%|███▎ | 9.96GB / 29.6GB, 872MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 34%|███▍ | 10.1GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 34%|███▍ | 10.1GB / 29.6GB, 872MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 35%|███▍ | 10.3GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 35%|███▍ | 10.3GB / 29.6GB, 874MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 35%|███▌ | 10.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 36%|███▌ | 10.5GB / 29.6GB, 876MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 36%|███▌ | 10.7GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 36%|███▌ | 10.7GB / 29.6GB, 877MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 37%|███▋ | 10.8GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 37%|███▋ | 10.9GB / 29.6GB, 876MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 37%|███▋ | 11.0GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 37%|███▋ | 11.0GB / 29.6GB, 876MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 38%|███▊ | 11.2GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 38%|███▊ | 11.2GB / 29.6GB, 876MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 38%|███▊ | 11.3GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 38%|███▊ | 11.4GB / 29.6GB, 874MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 39%|███▉ | 11.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 39%|███▉ | 11.6GB / 29.6GB, 876MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 40%|███▉ | 11.7GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 40%|███▉ | 11.7GB / 29.6GB, 879MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 40%|████ | 11.9GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 40%|████ | 11.9GB / 29.6GB, 879MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 41%|████ | 12.1GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 41%|████ | 12.1GB / 29.6GB, 881MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 41%|████▏ | 12.2GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 41%|████▏ | 12.3GB / 29.6GB, 881MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 42%|████▏ | 12.4GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 42%|████▏ | 12.4GB / 29.6GB, 882MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 42%|████▏ | 12.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 43%|████▎ | 12.6GB / 29.6GB, 881MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 43%|████▎ | 12.7GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 43%|████▎ | 12.7GB / 29.6GB, 882MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 44%|████▎ | 12.9GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 44%|████▎ | 12.9GB / 29.6GB, 884MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 44%|████▍ | 13.1GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 44%|████▍ | 13.1GB / 29.6GB, 885MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 45%|████▍ | 13.3GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 45%|████▍ | 13.3GB / 29.6GB, 885MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 46%|████▌ | 13.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 46%|████▌ | 13.5GB / 29.6GB, 893MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 46%|████▌ | 13.6GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 46%|████▌ | 13.7GB / 29.6GB, 895MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 47%|████▋ | 13.8GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 47%|████▋ | 13.9GB / 29.6GB, 896MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 47%|████▋ | 14.0GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 47%|████▋ | 14.0GB / 29.6GB, 898MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 48%|████▊ | 14.2GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 48%|████▊ | 14.2GB / 29.6GB, 899MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 49%|████▊ | 14.4GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 49%|████▊ | 14.4GB / 29.6GB, 900MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 49%|████▉ | 14.6GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 49%|████▉ | 14.6GB / 29.6GB, 901MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 50%|████▉ | 14.7GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 50%|████▉ | 14.8GB / 29.6GB, 901MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 51%|█████ | 14.9GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 51%|█████ | 15.0GB / 29.6GB, 901MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 51%|█████ | 15.1GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 51%|█████ | 15.1GB / 29.6GB, 901MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 52%|█████▏ | 15.3GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 52%|█████▏ | 15.3GB / 29.6GB, 901MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 52%|█████▏ | 15.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 52%|█████▏ | 15.5GB / 29.6GB, 900MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 53%|█████▎ | 15.6GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 53%|█████▎ | 15.7GB / 29.6GB, 900MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 53%|█████▎ | 15.8GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 54%|█████▎ | 15.8GB / 29.6GB, 897MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 54%|█████▍ | 16.0GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 54%|█████▍ | 16.0GB / 29.6GB, 897MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 55%|█████▍ | 16.1GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 55%|█████▍ | 16.2GB / 29.6GB, 896MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 55%|█████▌ | 16.3GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 55%|█████▌ | 16.4GB / 29.6GB, 896MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 56%|█████▌ | 16.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 56%|█████▌ | 16.5GB / 29.6GB, 896MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 56%|█████▋ | 16.7GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 57%|█████▋ | 16.7GB / 29.6GB, 896MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 57%|█████▋ | 16.9GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 57%|█████▋ | 16.9GB / 29.6GB, 896MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 58%|█████▊ | 17.0GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 58%|█████▊ | 17.1GB / 29.6GB, 896MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 58%|█████▊ | 17.2GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 58%|█████▊ | 17.3GB / 29.6GB, 895MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 59%|█████▉ | 17.4GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 59%|█████▉ | 17.4GB / 29.6GB, 893MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 59%|█████▉ | 17.6GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 60%|█████▉ | 17.6GB / 29.6GB, 893MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 60%|██████ | 17.8GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 60%|██████ | 17.8GB / 29.6GB, 893MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 61%|██████ | 17.9GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 61%|██████ | 18.0GB / 29.6GB, 893MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 61%|██████▏ | 18.1GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 61%|██████▏ | 18.2GB / 29.6GB, 894MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 62%|██████▏ | 18.3GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 62%|██████▏ | 18.3GB / 29.6GB, 894MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 63%|██████▎ | 18.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 63%|██████▎ | 18.5GB / 29.6GB, 894MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 63%|██████▎ | 18.7GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 63%|██████▎ | 18.7GB / 29.6GB, 894MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 64%|██████▍ | 18.9GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 64%|██████▍ | 18.9GB / 29.6GB, 894MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 64%|██████▍ | 19.0GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 64%|██████▍ | 19.1GB / 29.6GB, 891MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 65%|██████▍ | 19.2GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 65%|██████▌ | 19.2GB / 29.6GB, 891MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 66%|██████▌ | 19.4GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 66%|██████▌ | 19.4GB / 29.6GB, 890MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 66%|██████▌ | 19.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 66%|██████▌ | 19.6GB / 29.6GB, 889MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 67%|██████▋ | 19.7GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 67%|██████▋ | 19.8GB / 29.6GB, 889MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 67%|██████▋ | 19.9GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 67%|██████▋ | 19.9GB / 29.6GB, 890MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 68%|██████▊ | 20.1GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 68%|██████▊ | 20.1GB / 29.6GB, 889MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 69%|██████▊ | 20.3GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 69%|██████▊ | 20.3GB / 29.6GB, 889MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 69%|██████▉ | 20.4GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 69%|██████▉ | 20.5GB / 29.6GB, 891MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 70%|██████▉ | 20.6GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 70%|██████▉ | 20.6GB / 29.6GB, 889MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 70%|███████ | 20.8GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 70%|███████ | 20.8GB / 29.6GB, 889MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 71%|███████ | 21.0GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 71%|███████ | 21.0GB / 29.6GB, 889MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 72%|███████▏ | 21.1GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 72%|███████▏ | 21.2GB / 29.6GB, 889MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 72%|███████▏ | 21.3GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 72%|███████▏ | 21.3GB / 29.6GB, 891MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 73%|███████▎ | 21.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 73%|███████▎ | 21.5GB / 29.6GB, 893MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 73%|███████▎ | 21.7GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 73%|███████▎ | 21.7GB / 29.6GB, 894MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 74%|███████▍ | 21.8GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 74%|███████▍ | 21.9GB / 29.6GB, 895MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 75%|███████▍ | 22.0GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 75%|███████▍ | 22.1GB / 29.6GB, 894MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 75%|███████▌ | 22.2GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 75%|███████▌ | 22.2GB / 29.6GB, 893MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 76%|███████▌ | 22.4GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 76%|███████▌ | 22.4GB / 29.6GB, 893MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 76%|███████▋ | 22.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 76%|███████▋ | 22.6GB / 29.6GB, 890MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 77%|███████▋ | 22.7GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 77%|███████▋ | 22.7GB / 29.6GB, 889MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 77%|███████▋ | 22.9GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 77%|███████▋ | 22.9GB / 29.6GB, 889MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 78%|███████▊ | 23.1GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 78%|███████▊ | 23.1GB / 29.6GB, 888MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 79%|███████▊ | 23.2GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 79%|███████▊ | 23.3GB / 29.6GB, 887MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 79%|███████▉ | 23.4GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 79%|███████▉ | 23.4GB / 29.6GB, 885MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 80%|███████▉ | 23.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 80%|███████▉ | 23.6GB / 29.6GB, 882MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 80%|████████ | 23.7GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 80%|████████ | 23.8GB / 29.6GB, 881MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 81%|████████ | 23.9GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 81%|████████ | 23.9GB / 29.6GB, 878MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 81%|████████▏ | 24.1GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 81%|████████▏ | 24.1GB / 29.6GB, 878MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 82%|████████▏ | 24.2GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 82%|████████▏ | 24.3GB / 29.6GB, 878MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 83%|████████▎ | 24.4GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 83%|████████▎ | 24.4GB / 29.6GB, 878MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 83%|████████▎ | 24.6GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 83%|████████▎ | 24.6GB / 29.6GB, 876MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 84%|████████▍ | 24.7GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 84%|████████▍ | 24.8GB / 29.6GB, 878MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 84%|████████▍ | 24.9GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 84%|████████▍ | 25.0GB / 29.6GB, 878MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 85%|████████▍ | 25.1GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 85%|████████▍ | 25.1GB / 29.6GB, 876MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 86%|████████▌ | 25.3GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 86%|████████▌ | 25.3GB / 29.6GB, 876MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 86%|████████▌ | 25.4GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 86%|████████▌ | 25.4GB / 29.6GB, 874MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 87%|████████▋ | 25.6GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 87%|████████▋ | 25.6GB / 29.6GB, 872MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 87%|████████▋ | 25.7GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 87%|████████▋ | 25.8GB / 29.6GB, 871MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 88%|████████▊ | 25.9GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 88%|████████▊ | 26.0GB / 29.6GB, 871MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 88%|████████▊ | 26.1GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 88%|████████▊ | 26.1GB / 29.6GB, 868MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 89%|████████▉ | 26.3GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 89%|████████▉ | 26.3GB / 29.6GB, 869MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 89%|████████▉ | 26.4GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 89%|████████▉ | 26.4GB / 29.6GB, 866MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 90%|████████▉ | 26.6GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 90%|████████▉ | 26.6GB / 29.6GB, 865MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 91%|█████████ | 26.7GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 91%|█████████ | 26.8GB / 29.6GB, 862MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 91%|█████████ | 26.9GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 91%|█████████ | 26.9GB / 29.6GB, 860MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 92%|█████████▏| 27.1GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 92%|█████████▏| 27.1GB / 29.6GB, 859MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 92%|█████████▏| 27.2GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 92%|█████████▏| 27.3GB / 29.6GB, 857MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 93%|█████████▎| 27.4GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 93%|█████████▎| 27.4GB / 29.6GB, 856MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 93%|█████████▎| 27.6GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 93%|█████████▎| 27.6GB / 29.6GB, 853MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 94%|█████████▍| 27.7GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 94%|█████████▍| 27.8GB / 29.6GB, 854MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 94%|█████████▍| 27.9GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 94%|█████████▍| 27.9GB / 29.6GB, 853MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 95%|█████████▍| 28.1GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 95%|█████████▍| 28.1GB / 29.6GB, 851MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 96%|█████████▌| 28.2GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 96%|█████████▌| 28.3GB / 29.6GB, 851MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 96%|█████████▌| 28.4GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 96%|█████████▌| 28.4GB / 29.6GB, 852MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 97%|█████████▋| 28.6GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 97%|█████████▋| 28.6GB / 29.6GB, 852MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 97%|█████████▋| 28.8GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 97%|█████████▋| 28.8GB / 29.6GB, 852MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 98%|█████████▊| 28.9GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 98%|█████████▊| 29.0GB / 29.6GB, 852MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 99%|█████████▊| 29.1GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 99%|█████████▊| 29.2GB / 29.6GB, 853MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 99%|█████████▉| 29.3GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 99%|█████████▉| 29.3GB / 29.6GB, 853MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 100%|█████████▉| 29.5GB / 29.6GB, 853MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 100%|█████████▉| 29.6GB / 29.6GB, 840MB/s
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 100%|█████████▉| 29.6GB / 29.6GB, 788MB/s
|
||
New Data Upload : 5%|▍ | 608kB / 13.1MB, 59.6kB/s [A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 100%|█████████▉| 29.6GB / 29.6GB, 771MB/s
|
||
New Data Upload : 46%|████▋ | 6.08MB / 13.1MB, 596kB/s [A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (2 / 4) : 100%|█████████▉| 29.6GB / 29.6GB, 754MB/s
|
||
New Data Upload : 97%|█████████▋| 12.8MB / 13.1MB, 1.25MB/s [A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|█████████▉| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|█████████▉| 22.2MB / 22.2MB [A[A[A[A[A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (4 / 4) : 100%|██████████| 29.6GB / 29.6GB, 720MB/s
|
||
New Data Upload : 100%|██████████| 13.1MB / 13.1MB, 1.29MB/s [A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A[A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A[A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A[A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A[A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A[A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A[A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A[A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A[A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A[A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A[A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A[A
|
||
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB [A[A
|
||
|
||
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB [A[A[A
|
||
|
||
|
||
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB [A[A[A[A
|
||
|
||
|
||
|
||
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB [A[A[A[A[A
Processing Files (4 / 4) : 100%|██████████| 29.6GB / 29.6GB, 529MB/s
|
||
New Data Upload : 100%|██████████| 13.1MB / 13.1MB, 1.31MB/s
|
||
...ing/fft/training_args.bin: 100%|██████████| 7.38kB / 7.38kB
|
||
...coding/fft/tokenizer.json: 100%|██████████| 11.4MB / 11.4MB
|
||
...ing/fft/model.safetensors: 100%|██████████| 29.5GB / 29.5GB
|
||
...-30T11-12-27.892956.jsonl: 100%|██████████| 22.2MB / 22.2MB
|