Unleash the full potential of exascale LLMs on consumer-class GPUs, proven by extensive benchmarks, with no long-term adjustments and minimal learning curve.
Repositories
kyegomez repositories
This paper presents a comprehensive analysis of Kurt G¨odel’s ontological proof for the existence of God, with particular focus on contemporary expansions and refinements of the original framework
Fast and memory-efficient exact attention coupled with the LION Optimizer for ultra fast performance
An attempt to create the most accurate, reliable, and general vision transformers for facial recognition at scale.
FabLang is a domain-specific programming language crafted to streamline industrial design and make parts easier to manufacture.
Pretrained Facial Recognition Models to democratize access
A simple package for leveraging Falcon 180B and the HF ecosystem's tools, including training/inference scripts, safetensors, integrations with bitsandbytes, PEFT, GPTQ, assisted generation, RoPE scaling support, and rich generation parameters.
Zeta implementation of a reusable and plug in and play feedforward from the paper "Exponentially Faster Language Modeling"
Finetune any model on HF in less than 30 seconds
Flamingo Implemnted in Zeta + Pytorch primitives for high performance multi-modal learning
Get down and dirty with FlashAttention2.0 in pytorch, plug in and play no complex CUDA kernels
Triton implementation of Flash Attention2.0
FlashAttention2.0 with Lora
An simple pytorch implementation of Flash MultiHead Attention
An extremely experimental model that intakes images and generates 3D scenes of those images using Diffusion
Implementation of Adepts Fuyu all-new Multi-Modality model in pytorch
Unofficial GATO: Ready to train
Implementation of GATS from the paper: "GATS: Gather-Attend-Scatter" in pytorch and zeta
An implementation of the base GPT-3 Model architecture from the paper by OPENAI "Language Models are Few-Shot Learners"
The open source implementation of the base model behind GPT-4 from OPENAI [Language + Multi-Modal]