Integrating Mamba/SSMs with Transformer for Enhanced Long Context and High-Quality Sequence Modeling
Repositories
kyegomez repositories
All mathematics research papers sourced from ArXiv and meticulously curated for LLM pretraining purposes.
Democratization of Med-Flamingo, "Med-Flamingo: A Multimodal Medical Few-shot Learner"
Towards Generalist Biomedical AI
The open source implementation of the model from "Scaling Vision Transformers to 22 Billion Parameters"
Tree of Thoughts with an meta prompt for 50% boost in model reasoning
Implementation of Midas from [Towards Robust Monocular Depth Estimation] in Pytorch and Zeta
Democratization of "Cinematic Mindscapes: High-quality Video Reconstruction from Brain Activity" in Pytorch
Implementation of Minerva from "Minerva: Solving Quantitative Reasoning Problems with Language Models"
Pytorch Implementation of the Model from "MIRASOL3B: A MULTIMODAL AUTOREGRESSIVE MODEL FOR TIME-ALIGNED AND CONTEXTUAL MODALITIES"
Implementation of the paper: "Mixture-of-Depths: Dynamically allocating compute in transformer-based language models"
An implementation of a switch transformer like Multi-query attention model
Implementation of MoE Mamba from the paper: "MoE-Mamba: Efficient Selective State Space Models with Mixture of Experts" in Pytorch and Zeta
Implementation of the LDP module block in PyTorch and Zeta from the paper: "MobileVLM: A Fast, Strong and Open Vision Language Assistant for Mobile Devices"
A template for deploying ultra powerful LLMs effortlessly with the best optimizations
A simple, thorough, and reliable template I use to build new models from scratch:
The Ultimate Hub of AI Models - Simplified, Streamlined, and Scalable for Production-Grade Deployment.
This paper presents a groundbreaking framework that models economic systems as intelligent neural networks, offering a novel approach to understanding how economies learn, adapt, and self-organize.
An experimental inquiry into monte carlo tree of thoughts algorithmic systems
🔥 chat with over 10K frames of video!