Chess policy-value models and datasets.
👋 Open to Work
Andrew Thompson
AndrewThompson1233
P(doom) <0.001%
AI & ML interests
Deep Learning / Model Architecture & Training
Recent Activity
new activity 1 day ago
AndrewThompson1233/maba-v1-architecture:Factorized embeddings intervention for GPT-2 new activity 3 days ago
AndrewThompson1233/maba-instant-v1:Model? new activity 3 days ago
AwareLiquid/M2-2B:Horizon bottleneck at 128-byte sequence length and associative state updates