PatchWorld: Gradient-Free Optimization of Executable World Models Paper • 2605.30880 • Published May 29 • 12
MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models Paper • 2605.14906 • Published May 14 • 79
Training Long-Context Vision-Language Models Effectively with Generalization Beyond 128K Context Paper • 2605.13831 • Published May 13 • 90
zhongweixie/Llama-3.2-1B-NEW-dpo-math-5epoch-lr2e-7-bs16-beta0.01-entropy_non_linear Updated Jul 26, 2025
zhongweixie/Llama-3.2-1B-NEW-dpo-math-30epoch-lr2e-7-bs16-beta0.01-entropy_non_linear_15001 Updated Jul 1, 2025
zhongweixie/Llama-3.2-1B-NEW-dpo-math-30epoch-lr2e-7-bs16-beta0.01-entropy_non_linear_1500 Text Generation • 1B • Updated Jul 1, 2025 • 10
zhongweixie/Llama-3.2-1B-NEW-dpo-math-30epoch-lr2e-7-bs16-beta0.01-entropy_non_linear_1500 Text Generation • 1B • Updated Jul 1, 2025 • 10