Skip to content
View Concyclics's full-sized avatar
  • National University of Singapore
  • Singapore
  • 15:24 (UTC +08:00)

Block or report Concyclics

Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Accelerating MoE with IO and Tile-aware Optimizations

Python 775 108 Updated Aug 29, 2026

VideoSys: An easy and efficient system for video generation

Python 2,025 129 Updated Aug 27, 2025

高性价比人生指南: 长寿防病、急救、省钱理财、法律红线、失业与工伤、医保社保、恋爱婚育、怀孕育儿、创业与做平台合规、出国与技能。每条写明成本、收益、证据等级和原始出处,只引期刊论文与官方文件。

HTML 32,545 2,460 Updated Oct 1, 2026

HieraSparse: Hierarchical Semi-Structured KV-Cache Attention on Sparse Tensor Core

Cuda 9 1 Updated Sep 15, 2026

Code for the paper “SharQ: Bridging Activation Sparsity and FP4 Quantization for LLM Inference”

Python 15 2 Updated Jul 1, 2026

Code for the paper "Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning" https://arxiv.org/abs/2609.03430

Python 71 9 Updated Sep 4, 2026

My Python scripts to make high-quality figures for publications in top AI conferences and journals.

Python 7,889 548 Updated Sep 26, 2026

Defend sound research against LLM reviewer bias through meaning-preserving rewrites.

200 5 Updated Sep 22, 2026

Benchmarking Knowledge Transfer in Lifelong Robot Learning

Jupyter Notebook 2,373 507 Updated Mar 15, 2025

Official repository of LIBERO-plus, a generalized benchmark for in-depth robustness analysis of vision-language-action models.

Python 469 51 Updated Jan 21, 2026

Measuring and evolving with the frontier of agent work

Python 823 550 Updated Oct 1, 2026

A Vulkan renderer for side-by-side quality and performance comparison of upscaling algorithms.

C 65 9 Updated Sep 7, 2026

Memory is a long-term memory module for AI agents, providing capabilities for memory extraction, storage, retrieval, and migration for agents running on the openJiuwen framework.

Python 64 19 Updated Sep 22, 2026

SWE-Bench Pro: Can AI Agents Solve Long-Horizon Software Engineering Tasks?

Python 534 101 Updated Sep 22, 2026

Fast, Flexible and Portable Structured Generation

C++ 1 Updated Sep 4, 2026

FlashInfer: Kernel Library for LLM Serving

Python 1 Updated Sep 4, 2026
Python 6 Updated Sep 11, 2026

Benchmarking physical understanding in generative video models

Python 355 49 Updated Oct 1, 2026

KaHIP -- HIGH Quality Partitioning.

C++ 499 107 Updated Sep 16, 2026

[NeurIPS'24] HippoRAG is a novel RAG framework inspired by human long-term memory that enables LLMs to continuously integrate knowledge across external documents. RAG + Knowledge Graphs + Personali…

Python 4,034 434 Updated Sep 29, 2026

From Foundation to Application

Python 991 114 Updated Sep 22, 2026

PhyAI is a high-performance framework for running Physical AI models (VLA, WAM, and beyond), supporting both cloud-based serving and on-device deployment.

Python 127 27 Updated Sep 29, 2026

Sub2API 一站式开源中转服务,让 Claude、Openai 、Gemini、Grok订阅统一接入,支持拼车共享,更高效分摊成本,原生工具无缝使用。

Go 43,153 9,227 Updated Sep 30, 2026

PonderTTT: Adaptive Budget-Aware Test-Time Training

Python 9 2 Updated Feb 6, 2026

TokenSpeed is a speed-of-light LLM inference engine.

Python 2,187 299 Updated Oct 1, 2026

XtraMAC code repo (Accepted by ISCA2026)

Verilog 31 5 Updated Sep 9, 2026

Graphs for Everyone

Java 17,269 2,703 Updated Sep 22, 2026

Milvus is a high-performance, cloud-native vector database built for scalable vector ANN search

Go 46,294 4,279 Updated Sep 30, 2026

提供一个人人会用的的路由、NAS系统 (目前活跃的分支是 istoreos-24.10,main或master分支不维护请勿使用)

C 8,119 928 Updated Sep 24, 2026

Official implementation of "Fast-dLLM: Training-free Acceleration of Diffusion LLM by Enabling KV Cache and Parallel Decoding"

Python 1,093 145 Updated May 30, 2026
Next