Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

WatchingResearchWatching0 independent reports0

Probing Latent Colombian Identity Inferences in Qwen2.5-7B with Natural Language Autoencoders

First seen · 7/24/2026, 03:42 AMLatest activity · 7/24/2026, 03:42 AM

This pilot study uses Natural Language Autoencoders to verbalize layer-20 residual-stream activations in Qwen2.5-7B-Instruct. It examines 30 prompts organized as 15 matched Colombian-Spanish and English pairs, covering explicit Colombian cues, implicit cues, and neutral controls. Activations are sampled across four positional quartiles to investigate whether nationality, socioeconomic status, or stereotype-related information appears internally before reaching the model output. The authors report descriptive rates and qualitative observations rather than statistically powered effects, so the work should be read as an exploratory interpretability and bias-evaluation study.

Event heat · last 24 hours

No heat snapshots are available in the last 24 hours.

No heat snapshots are available in the last 24 hours.

Reporting Timeline

  1. AggregatorarXiv7/24, 03:42 AMnot independentRepresentative
    Probing Latent Colombian Identity Inferences in Qwen2.5-7B with Natural Language Autoencoders