Hacker Newsnew | past | comments | ask | show | jobs | submit | fromlogin
XGrammar-2: Fast, Customizable Structured Generation for Tool Calling and Agents (mlc.ai)
2 points by matt_d 5 days ago | past | discuss
Modern GPU Programming for MLSys (mlc.ai)
2 points by sonabinu 3 months ago | past | 1 comment
Modern GPU Programming Book (mlc.ai)
2 points by fuy 3 months ago | past | 1 comment
Modern GPU Programming for MLSys Book (mlc.ai)
1 point by tanelpoder 3 months ago | past
Modern GPU Programming for MLSys (mlc.ai)
80 points by crowwork 3 months ago | past | 25 comments
PithTrain – a compact, agent-native MoE training system (mlc.ai)
4 points by ruihangl 3 months ago | past
XGrammar-2: 80x Faster Structured Generation for Agent Tool Calling (mlc.ai)
6 points by ubospica 4 months ago | past
Machine Learning Compiler (mlc.ai)
1 point by selvan on July 24, 2025 | past
Microserving LLM Engines (mlc.ai)
1 point by homarp on Jan 13, 2025 | past | 1 comment
LLM Microserving: a new RISC-style approach to design LLM serving API (mlc.ai)
4 points by jinhongyii on Jan 7, 2025 | past | 1 comment
Making AMD GPUs competitive for LLM inference (2023) (mlc.ai)
313 points by plasticchris on Dec 24, 2024 | past | 213 comments
Optimizing and Characterizing High-Throughput Low-Latency LLM Inference (mlc.ai)
1 point by djhu9 on Oct 11, 2024 | past
High-Throughput Low-Latency LLM Serving with MLCEngine (mlc.ai)
8 points by ruihangl on Oct 10, 2024 | past
In-browser LLM inference engine with WebGPU and OpenAI API (mlc.ai)
16 points by CharlieRuan on June 13, 2024 | past | 4 comments
MLCEngine: Universal LLM Deployment to Both Cloud and Local Devices (mlc.ai)
2 points by crowwork on June 8, 2024 | past
Universal LLM Deployment Engine with ML Compilation (mlc.ai)
17 points by ruihangl on June 7, 2024 | past | 7 comments
MLC LLM: Universal Language Model Deployment Across Diverse Hardware and Apps (mlc.ai)
1 point by georgehill on Dec 16, 2023 | past
Scaling LLama2-70B with Multiple Nvidia/AMD GPU (mlc.ai)
13 points by junrushao1994 on Oct 20, 2023 | past | 6 comments
WebLLM: Llama2 in the Browser (mlc.ai)
192 points by meiraleal on Aug 29, 2023 | past | 31 comments
GPU-Accelerated LLM on an Orange Pi (mlc.ai)
214 points by tosh on Aug 15, 2023 | past | 80 comments
Making AMD GPUs competitive for LLM inference (mlc.ai)
354 points by djoldman on Aug 9, 2023 | past | 132 comments
Run Llama2-70B in Web Browser with WebGPU Acceleration (mlc.ai)
9 points by ruihangl on July 24, 2023 | past | 6 comments
Bringing Open Large Language Models to Consumer Devices (mlc.ai)
31 points by hardmaru on May 23, 2023 | past
Running RedPajama and other open LLMs on phones, browsers and AMD/NV/Intel GPUs (mlc.ai)
11 points by junrushao1994 on May 23, 2023 | past
Bringing Open Large Language Models to Consumer Devices (mlc.ai)
11 points by shantanu_sharma on May 22, 2023 | past
Browser-based Stable Diffusion using WebGPU (mlc.ai)
3 points by Eduard on May 6, 2023 | past
Bringing Hardware Accelerated Language Models to Consumer Devices (mlc.ai)
1 point by crowwork on May 1, 2023 | past
MLC: Bringing Hardware Accelerated Language Models to Consumer Devices (mlc.ai)
8 points by junrushao1994 on May 1, 2023 | past
What Is ML Compilation (mlc.ai)
88 points by tosh on April 30, 2023 | past | 5 comments
Vicuna on iPhone (mlc.ai)
90 points by tosh on April 30, 2023 | past | 15 comments

Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: