Skip to content
View bchao1's full-sized avatar
🚶‍♂️
I need to focus.
🚶‍♂️
I need to focus.

Block or report bchao1

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

A unified multimodal model toolkit

Python 582 114 Updated Aug 12, 2026

[ICLR & NeurIPS 2025] Repository for Show-o series, One Single Transformer to Unify Multimodal Understanding and Generation.

Python 1,971 93 Updated Jan 8, 2026

An open-source AI agent that brings the power of Gemini directly into your terminal.

TypeScript 106,534 14,443 Updated Aug 16, 2026

[ICLR 2026] Official Implementation of Muddit [Meissonic II]: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model.

Python 121 2 Updated Apr 13, 2026

MMaDA - Open-Sourced Multimodal Large Diffusion Language Models (dLLMs with block diffusion, mixed-CoT, unified RL)

Python 1,663 91 Updated Feb 14, 2026

NVIDIA FastGen: Fast Generation from Diffusion Models

Python 944 80 Updated Aug 4, 2026

🪨 why use many token when few token do trick — Claude Code skill that cuts 65% of tokens by talking like caveman

Go 98,543 5,703 Updated Aug 16, 2026

Official implementation of "Phase-Aligned RoPE for Mixed-Resolution Diffusion Transformer" (ECCV 2026)

Python 4 Updated Jun 21, 2026

The official code for NeurIPS 2025 "MagCache: Fast Video Generation with Magnitude-Aware Cache"

Python 276 7 Updated Nov 17, 2025

LLM Wiki is a cross-platform desktop application that turns your documents into an organized, interlinked knowledge base — automatically. Instead of traditional RAG (retrieve-and-answer from scratc…

TypeScript 16,429 1,947 Updated Aug 14, 2026

Florence-2 is a novel vision foundation model with a unified, prompt-based representation for a variety of computer vision and vision-language tasks.

Jupyter Notebook 207 19 Updated Jul 3, 2024

Gemma open-weight LLM library, from Google DeepMind

Python 5,659 1,012 Updated Aug 10, 2026
Python 112 Updated Jun 12, 2026

Code repository for "Spectral Progressive Diffusion for Efficient Image and Video Generation"

Python 15 2 Updated Aug 13, 2026

Code release for "Foveated Diffusion: Efficient Spatially Adaptive Image and Video Generation"

Python 18 3 Updated Aug 2, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 2 Updated Aug 10, 2026

An agentic skills framework & software development methodology that works.

Shell 272,795 24,389 Updated Aug 13, 2026

Ideogram 4: Open image model at the forefront of design

Python 2,734 280 Updated Jun 30, 2026

Wrapper of 50+ image matching models with a unified interface

Python 902 79 Updated Jul 11, 2026

Implementation of Gamma-World: Generative Multi-Agent World Modeling Beyond Two Players

Python 654 13 Updated Jun 17, 2026

A PyTorch-native inference engine with cache, parallelism, quantization and cpu offload for DiTs.

Python 1,248 77 Updated Aug 15, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 31,920 7,933 Updated Aug 16, 2026

ComfyUI Unnofficial Implementation of Spectral Progressive Diffusion for Efficient Image and Video Generation for Anima

Python 70 9 Updated Jul 4, 2026

GDM Science Skills to speed up agentic scientific workflows with better grounding and higher token efficiency. Integrate insights from AlphaGenome, AFDB, UniProt and 30+ other databases and tools.

Python 2,710 294 Updated Jul 7, 2026

CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image

Jupyter Notebook 34,178 4,042 Updated Mar 25, 2026

Algorithm powering the For You feed on X

Rust 31,491 5,164 Updated Aug 14, 2026

Efficient PyTorch Hessian eigendecomposition tools!

Python 474 49 Updated Aug 9, 2026

[CVPR 2026 Highlight] A Frame is Worth One Token: Efficient Generative World Modeling with Delta Tokens

Python 231 8 Updated Jul 17, 2026

Official repository for “PixelGen: Improving Pixel Diffusion with Perceptual Loss”

Python 276 14 Updated May 12, 2026

[CVPR 2026] Denoising, Fast and Slow: Difficulty-Aware Adaptive Sampling for Image Generation

Python 102 4 Updated Apr 26, 2026
Next