ReactiveAI 's Collections

Sparse Query Attention (SQA) Research

Experimental models with Sparse Query Attention layers. Reducing training time/cost by ~3-10% compared to GQA & MQA, with the same level performance

This collection has no items.