Skip to content
View whwu95's full-sized avatar
♥️
I may be slow to respond.
♥️
I may be slow to respond.

Highlights

  • Pro

Block or report whwu95

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Create beautiful slides on the web using a coding agent's frontend skills

JavaScript 27,621 2,240 Updated Jun 23, 2026

An open-source implementaion for fine-tuning Qwen-VL series by Alibaba Cloud.

Python 1,959 221 Updated Jul 25, 2026

Wan: Open and Advanced Large-Scale Video Generative Models

Python 16,836 3,378 Updated Mar 5, 2026

EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL

Python 5,117 384 Updated Jul 30, 2026

A Scientific Multimodal Foundation Model

845 47 Updated Jul 17, 2026

[CVPR 2025 Oral] VideoEspresso: A Large-Scale Chain-of-Thought Dataset for Fine-Grained Video Reasoning via Core Frame Selection

Python 141 4 Updated Jul 28, 2025

Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL

Python 5,297 479 Updated Nov 13, 2025

Awesome Reasoning in MLLMs: Papers and Projects about learning to reason with MLLMs, including Chain-of-Thought (CoT), OpenAl o1, and DeepSeek-R1

63 4 Updated Mar 18, 2025
TeX 123 54 Updated Jan 29, 2025

[NIPS'25 Spotlight] Mulberry, an o1-like Reasoning and Reflection MLLM Implemented via Collective MCTS

Python 1,243 113 Updated Jan 16, 2026

Efficient Multimodal Large Language Models: A Survey

387 21 Updated Apr 29, 2025

Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.

Jupyter Notebook 19,793 1,830 Updated Jan 30, 2026

A series of math-specific large language models of our Qwen2 series.

Python 1,082 161 Updated Jan 11, 2025

LMDeploy is a toolkit for compressing, deploying, and serving LLMs.

Python 8,009 724 Updated Aug 14, 2026

The official repo of Qwen2-Audio chat & pretrained large audio language model proposed by Alibaba Cloud.

Python 2,098 168 Updated Apr 21, 2025

Retrieval-Augmented Generation in 3 Lines of Code!

Python 56 10 Updated Aug 14, 2026

AudioBench: A Universal Benchmark for Audio Large Language Models

Python 323 17 Updated May 29, 2026

【NeurIPS 2024】The official code of paper "Automated Multi-level Preference for MLLMs"

Python 22 1 Updated Sep 26, 2024

【NeurIPS 2024】Dense Connector for MLLMs

Python 182 8 Updated Oct 14, 2024

FreeVA: Offline MLLM as Training-Free Video Assistant

Python 69 1 Updated Jun 9, 2024

AcadHomepage: A Modern and Responsive Academic Personal Homepage

SCSS 2,900 5,820 Updated Aug 14, 2026

Awesome-LLM-Tabular: a curated list of Large Language Model applied to Tabular Data

428 35 Updated Apr 6, 2026

GPT4Vis: What Can GPT-4 Do for Zero-shot Visual Recognition?

Python 184 18 Updated May 22, 2024

【ICCV'2023】What Can Simple Arithmetic Operations Do for Temporal Modeling?

Python 74 6 Updated Jan 26, 2024

Demonstrate all the questions on LeetCode in the form of animation.(用动画的形式呈现解LeetCode题目的思路,完整单步/回看/变速/语音讲解在 algomooc.com)

Java 76,693 13,885 Updated Jun 12, 2026

Enjoy https://shields.io

Go 461 245 Updated Aug 2, 2026
JavaScript 4,328 1,938 Updated Jun 21, 2024

[ICCV 2023] Official Implementation of "Generalized Lightness Adaptation with Channel Selective Normalization"

Python 86 6 Updated Jan 22, 2024

A curated list of papers and open-source resources focused on 3D AIGC.

348 25 Updated Jul 2, 2026

The largest curated collection of markdown badges for your personal developer branding, profile, and projects.

SCSS 16,950 1,794 Updated Aug 11, 2026
Next