PFaithBench resistz/PairContext-Qwen3-4B 4B • Updated 3 days ago • 16 resistz/PFaithBench Viewer • Updated 3 days ago • 3.34k • 39 resistz/PairContext-Llama-3.2-3B-Instruct 4B • Updated 3 days ago • 15 resistz/PairContext-Llama-3.1-8B-Instruct 8B • Updated 3 days ago • 13
SFT on UltraChat200K Train one epoch SFT on UltraChat200K resistz/sft_Llama-3.1-8B_ultra200k_merged 8B • Updated Aug 13, 2025 • 9 resistz/sft_Qwen3-1.7B-Base_ultra200k Text Generation • 0.4B • Updated Aug 19, 2025 • 16 resistz/sft_Qwen3-8B-Base_ultra200k_lora32 Text Generation • Updated Aug 19, 2025 • 13 resistz/sft_Llama-3.2-1B_ultra200k Text Generation • 0.3B • Updated Aug 19, 2025 • 13
Test-time Calibration Learning (TTCL) Models Test-time calibration learning for large language model reasoning resistz/TTCL-MATH500-Qwen3-4B 4B • Updated 10 days ago • 22 resistz/TTCL-SimpleQA-Qwen3-8B 8B • Updated 10 days ago • 16 resistz/TTCL-AMC23-Qwen3-4B 4B • Updated 10 days ago • 27 resistz/TTCL-AIME24-Qwen3-4B 4B • Updated 10 days ago • 21
PFaithBench resistz/PairContext-Qwen3-4B 4B • Updated 3 days ago • 16 resistz/PFaithBench Viewer • Updated 3 days ago • 3.34k • 39 resistz/PairContext-Llama-3.2-3B-Instruct 4B • Updated 3 days ago • 15 resistz/PairContext-Llama-3.1-8B-Instruct 8B • Updated 3 days ago • 13
Test-time Calibration Learning (TTCL) Models Test-time calibration learning for large language model reasoning resistz/TTCL-MATH500-Qwen3-4B 4B • Updated 10 days ago • 22 resistz/TTCL-SimpleQA-Qwen3-8B 8B • Updated 10 days ago • 16 resistz/TTCL-AMC23-Qwen3-4B 4B • Updated 10 days ago • 27 resistz/TTCL-AIME24-Qwen3-4B 4B • Updated 10 days ago • 21
SFT on UltraChat200K Train one epoch SFT on UltraChat200K resistz/sft_Llama-3.1-8B_ultra200k_merged 8B • Updated Aug 13, 2025 • 9 resistz/sft_Qwen3-1.7B-Base_ultra200k Text Generation • 0.4B • Updated Aug 19, 2025 • 16 resistz/sft_Qwen3-8B-Base_ultra200k_lora32 Text Generation • Updated Aug 19, 2025 • 13 resistz/sft_Llama-3.2-1B_ultra200k Text Generation • 0.3B • Updated Aug 19, 2025 • 13