-
HUST
- Wuhan
-
19:05
(UTC +08:00)
Lists (1)
Sort Name ascending (A-Z)
Starred repositories
Curated list of project-based tutorials
Python tool for converting files and office documents to Markdown.
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Claude Code is an agentic coding tool that lives in your terminal, understands your codebase, and helps you code faster by executing routine tasks, explaining complex code, and handling git workflo…
Tensors and Dynamic neural networks in Python with strong GPU acceleration
A high-throughput and memory-efficient inference and serving engine for LLMs
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
为GPT/GLM等LLM大语言模型提供实用化交互接口,特别优化论文阅读/润色/写作体验,模块化设计,支持自定义快��按钮&函数插件,支持Python和C++等项目剖析&自译解功能,PDF/LaTex论文翻译&总结功能,支持并行问询多种LLM模型,支持chatglm3等本地模型。接入通义千问, deepseekcoder, 讯飞星火, 文心一言, llama2, rwkv, claude2, m…
Unsloth is a local UI for training and running Kimi K3, Gemma 4, Qwen3.6, DeepSeek-V4, GLM and other models.
The simplest, fastest repository for training/finetuning medium-sized GPTs.
LlamaIndex is the leading document agent and OCR platform
[EMNLP2025] "LightRAG: Simple and Fast Retrieval-Augmented Generation"
A modular graph-based Retrieval-Augmented Generation (RAG) system
SGLang is a high-performance serving framework for large language models and multimodal models.
Fast and memory-efficient exact attention
The official Python SDK for Model Context Protocol servers and clients
A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations
🚀 One-stop solution for creating your AI twin from chat history 💡 Fine-tune LLMs with your chat logs to capture your unique style, then bind to a chatbot to bring your digital self to life.
Agent framework and applications built upon Qwen>=3.0, featuring Function Calling, MCP, Code Interpreter, RAG, Chrome extension, etc.
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…
The official GitHub page for the survey paper "A Survey of Large Language Models".
Domain-specific language designed to streamline the development of high-performance GPU/CPU/Accelerators kernels
DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms
Genetic Algorithm, Particle Swarm Optimization, Simulated Annealing, Ant Colony Optimization Algorithm,Immune Algorithm, Artificial Fish Swarm Algorithm, Differential Evolution and TSP(Traveling sa…
FlashInfer: Kernel Library for LLM Serving
A framework for efficient model inference with omni-modality models


