GUI-HARVEST: Self-Improving GUI Agents through Evidence-Driven Harness Evolution Paper • 2610.00948 • Published 10 days ago • 19
Heoni/llama-3-KoEn-8b_sft_ep2_merged_red_teaming_20240614 Text Generation • Updated Jun 16, 2024 • 31 • 5
HLA: Expressive Hybrid Linear Attention via Chunk-Wise Dynamic Mixing Paper • 2610.05842 • Published 6 days ago • 9
andersonbcdefg/red_teaming_reward_modeling_pairwise_no_as_an_ai Viewer • Updated Jun 1, 2023 • 35.3k • 271 • 13
LoGRA: Scaling LLM Reinforcement Learning with Low-Rank Gradient Sketches Paper • 2610.06647 • Published 6 days ago • 92
Where Does Retrieval-Based Open-Ended Evaluation Fail? Automatic Taxonomy Induction from Long-Form Medical Answer Factuality Verification Paper • 2609.30467 • Published 17 days ago • 23
Triadic Linear Attention: Three-Dimensional Recurrent States for Long-Context Sequence Modeling Paper • 2609.36529 • Published 12 days ago • 37
InterEvolve: Test-Time Evolution of Reward Programs for Humanoid Loco-Manipulation Paper • 2610.02196 • Published 10 days ago • 62
Mindgard/evaded-prompt-injection-and-jailbreak-samples Viewer • Updated Apr 30, 2025 • 11.3k • 243 • 24