WHALE: A Simple Recipe for Joint Harness-Weight Optimization Paper • 2609.00196 • Published 9 days ago • 32
Set Transformer: A Framework for Attention-based Permutation-Invariant Neural Networks Paper • 1810.00825 • Published Oct 1, 2018 • 1
RLAD: Training LLMs to Discover Abstractions for Solving Reasoning Problems Paper • 2510.02263 • Published Oct 2, 2025 • 9
Discrete Infomax Codes for Supervised Representation Learning Paper • 1905.11656 • Published May 28, 2019
Learning Dynamics of Attention: Human Prior for Interpretable Machine Reasoning Paper • 1905.11666 • Published May 28, 2019
DetectGPT: Zero-Shot Machine-Generated Text Detection using Probability Curvature Paper • 2301.11305 • Published Jan 26, 2023 • 2
Self-Guided Masked Autoencoders for Domain-Agnostic Self-Supervised Learning Paper • 2402.14789 • Published Feb 22, 2024
WHALE: A Simple Recipe for Joint Harness-Weight Optimization Paper • 2609.00196 • Published 9 days ago • 32