arxiv:2608.08722
Victor Gallego
vicgalle
AI & ML interests
Preference fine-tuning, alignment & synthetic data.
Building LLMs in general!
Recent Activity
authored a paper 3 days ago
Opponent Aware Reinforcement Learning upvoted a paper 4 days ago
Opponent Aware Reinforcement Learning authored a paper 6 days ago
Gaming Without an Attacker: Benchmark Fingerprinting in LLM-Driven Search Under Selection Pressure