Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

Frustratingly Simple Black-Box Adaptation of Language Models via Logit Bias

First seen · 7/25/2026, 02:27 AMLatest activity · 7/25/2026, 02:27 AM

This paper studies a minimal black-box adaptation method that learns a single context-independent logit-bias vector and adds it at every decoding step. The intervention requires no weight updates or gradients. Starting from a KL-regularized reinforcement-learning objective, the authors characterize when a fixed bias can approximate an optimal prefix-dependent correction and derive a closed-form inverse-propensity estimator using rollouts, rewards, and token probabilities. The abstract reports gains over base models on mathematical and reasoning benchmarks, with substantially fewer trainable parameters than conventional fine-tuning.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/25, 02:27 AMnot independentRepresentative
    Frustratingly Simple Black-Box Adaptation of Language Models via Logit Bias