PPaperPicks

Piotr Milos

10 papers at tracked venues · 10 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. Contrastive Representations for Temporal Reasoning
  2. Joint MoE Scaling Laws: Mixture of Experts Can Be Memory Efficient
  3. Since Faithfulness Fails: The Performance Limits of Neural Causal Discovery
  4. Structured Packing in LLM Training Improves Long Context Utilization
  5. Analysing The Impact of Sequence Composition on Language Model Pre-Training
  6. Bigger, Regularized, Optimistic: scaling for compute and sample efficient continuous control
  7. Fine-tuning Reinforcement Learning Models is Secretly a Forgetting Mitigation Problem
  8. Magnushammer: A Transformer-Based Approach to Premise Selection
  9. Overestimation, Overfitting, and Plasticity in Actor-Critic: the Bitter Lesson of Reinforcement Learning
  10. Repurposing Language Models into Embedding Models: Finding the Compute-Optimal Recipe