r/MachineLearning
7/16/2026
![Seeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]](https://preview.redd.it/b0u6q9a46ndh1.jpg?width=140&height=98&auto=webp&s=dbb02d2e0fc85305a04e37864167c2578891d46c)
Seeking collaborators for scaling and independent evaluation of a new recurrent language model architecture (preprint + code) [R]
Short summary
An independent researcher shares a preprint and open-source code for DABSN (Dynamic Adaptive Bias State Network), a recurrent architecture evaluated on reasoning, memory, and long-sequence benchmarks. Initial language modeling results with a 24M parameter model trained on 1B tokens were promising enough to warrant a second paper. The author seeks collaborators for independent reproduction, stronger baselines, and access to larger GPU clusters for scaling.
- •DABSN is a novel recurrent architecture with PyTorch, C++, and Triton implementations, fully open-source
- •Initial 24M parameter language model trained on 1B tokens showed unexpectedly interesting results
- •Author seeks collaborators for reproduction, evaluation design, and GPU cluster access for scaling experiments
Generated with AI, which can make mistakes.
Is this a good recommendation for you?



