GRIDINDEX TOPIC

Tuning

22 tracked items

Hugging Face·8d ago
Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps
RESEARCHPhysical Review Letters·9d ago
Electronic Tuning of the Soft-Phonon Transport Anomaly in
RESEARCHarXiv·9d ago
A Glance Is All You Need: Single-Pass Fine-Grained Image Captioning with SimLoss
RESEARCHarXiv·9d ago
WHALE: A Simple Recipe for Joint Harness-Weight Optimization
RESEARCHarXiv·9d ago
Beyond Language Priors: Diagnosing and Fixing Visual-Origin Hallucinations in Multimodal LLM
RESEARCHarXiv·9d ago
AnyWorld: Factorized Egocentric World Models for Cross-Embodiment Generalization
RESEARCHarXiv·9d ago
Flawed in Nature, Perfect through Evolution
RESEARCHarXiv·9d ago
Safin-1: Safety from Within through Memory-Native State Evolution
RESEARCHarXiv·9d ago
DREAM: Deployment-Time Demonstration Generation via Real-to-Sim for Scalable Policy Adaptation
RESEARCHarXiv·9d ago
Elite-Weighted Supervised Fine-tuning for Goal-Directed Molecular Optimization
RESEARCHarXiv·9d ago
Uncovering and Mitigating Aggregation-Induced Reward Hacking in Multi-Reward Reinforcement Learning
RESEARCHarXiv·9d ago
LLM-as-a-Demographic: Whom Sociodemographic Prompting Helps, and Whom It Hurts
RESEARCHarXiv·9d ago
Synthetic Worlds for Temporal Evaluation and Knowledge Updating in LLMs
RESEARCHarXiv·9d ago
AdaVLA: Adaptive Step Flow Matching for Training-free Acceleration of Vision-Language-Action Models
RESEARCHarXiv·9d ago
A Cone-Constrained Bilinear Decomposition for Total Scaled-Gradient Variation Models
RESEARCHarXiv·9d ago
Bridging Lexical Divergence: LLM-Assisted, Cost-Efficient, Zero-shot Scientific Entity Linking
RESEARCHarXiv·9d ago
SAM3-LoRA: Parameter-Efficient Adaptation of a Concept-Promptable Foundation Model for Multi-Class Structural Defect Segmentation
RESEARCHarXiv·9d ago
Behaviorally Grounded User Profiles from the Wild for Personalized Alignment and Multi-Perspective Reasoning
RESEARCHarXiv·9d ago
Attention Sensitivity Is Not Enough: Dissociating Attention-Level and Behavioural In-Context Learning under Fine-Tuning
RESEARCHarXiv·9d ago
RW-LoRA: Communication-Efficient Decentralized LoRA Fine-Tuning via Random Walks
RESEARCHarXiv·9d ago
Do General NLP Embeddings Capture Ontological Reasoning?
RESEARCHbioRxiv·10d ago
Motor planning and execution establish distinct feedforward and feedback motor histories