Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

SyRuP: Enhancing System-Prompt Following via Reward-Guided Prediction in LLM Decoding

First seen · 7/27/2026, 12:39 PMLatest activity · 7/27/2026, 12:39 PM

SyRuP is an inference-time framework for improving system-prompt adherence while keeping the base language model frozen. It trains a cross-attention reward head on system-prompt-conditioned preference pairs, using the system prompt as a separate memory to score candidate tokens. During decoding, the method reranks the base model’s top-k candidates by combining base logits with the learned adherence reward and an optional contrastive signal based on system-induced logit shifts. The abstract reports consistent gains over prompting and decoding-time baselines with moderate inference overhead, but provides no detailed model, dataset, or numerical results.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/27, 12:39 PMnot independentRepresentative
    SyRuP: Enhancing System-Prompt Following via Reward-Guided Prediction in LLM Decoding