Hi everyone,
I’m an independent undergraduate starting my research journey in mechanistic interpretability.
Over the past couple of months I’ve completed my first research manuscript, which studies computational organization in a small character-level transformer through convergent qualitative evidence from attention circuits, hidden-state dynamics and cross-method observations.
The manuscript has gone through multiple rounds of revision, restructuring and external feedback, and I’m now preparing its first arXiv submission.
Since this is my first submission to cs.LG, arXiv requires an endorsement before submission.
If anyone here is an eligible endorser and would be willing to skim the manuscript (or read it more thoroughly) and, if they believe it is appropriate for cs.LG, consider endorsing it, I would be extremely grateful. I am not asking for a blind endorsement, only for someone willing to evaluate whether the work belongs in the category.
PDF:
I’d be happy to share the paper’s pdf privately if you’re interested
Thank you!
Aditya Pratap Singh