Skip to content
@vectorch-ai

Vectorch AI

Empowering Teams to Unleash Discoverability

Pinned Loading

  1. ScaleLLM ScaleLLM Public archive

    A high-performance inference system for large language models, designed for production environments.

    C++ 499 41

Repositories

Showing 10 of 16 repositories
  • nv-embedding-cache Public Forked from NVIDIA/nv-embedding-cache

    Fast hierarchical embedding cache for recommenders

    vectorch-ai/nv-embedding-cache's past year of commit activity
    C++ 0 Apache-2.0 13 0 0 Updated Jul 5, 2026
  • hpc-ops Public Forked from Tencent/hpc-ops

    High Performance LLM Inference Operator Library

    vectorch-ai/hpc-ops's past year of commit activity
    C++ 0 144 0 0 Updated Jun 9, 2026
  • ScaleLLM Public archive

    A high-performance inference system for large language models, designed for production environments.

    vectorch-ai/ScaleLLM's past year of commit activity
    C++ 499 Apache-2.0 41 48 (1 issue needs help) 8 Updated Dec 19, 2025
  • cutlass Public Forked from NVIDIA/cutlass

    CUDA Templates for Linear Algebra Subroutines

    vectorch-ai/cutlass's past year of commit activity
    C++ 0 2,127 0 0 Updated Nov 6, 2025
  • nixl Public Forked from ai-dynamo/nixl

    NVIDIA Inference Xfer Library (NIXL)

    vectorch-ai/nixl's past year of commit activity
    C++ 0 Apache-2.0 461 0 0 Updated Nov 4, 2025
  • whl Public

    repository to host python whl package.

    vectorch-ai/whl's past year of commit activity
    HTML 0 0 0 0 Updated Sep 13, 2025
  • flux Public Forked from bytedance/flux

    A fast communication-overlapping library for tensor/expert parallelism on GPUs.

    vectorch-ai/flux's past year of commit activity
    C++ 0 Apache-2.0 117 0 0 Updated Apr 15, 2025
  • 3FS Public Forked from deepseek-ai/3FS

    A high-performance distributed file system designed to address the challenges of AI training and inference workloads.

    vectorch-ai/3FS's past year of commit activity
    C++ 0 MIT 1,114 0 0 Updated Feb 28, 2025
  • flashinfer Public Forked from flashinfer-ai/flashinfer

    FlashInfer: Kernel Library for LLM Serving

    vectorch-ai/flashinfer's past year of commit activity
    Cuda 0 Apache-2.0 1,524 0 0 Updated Feb 27, 2025
  • FlashMLA Public Forked from deepseek-ai/FlashMLA
    vectorch-ai/FlashMLA's past year of commit activity
    C++ 0 MIT 1,190 0 0 Updated Feb 26, 2025

People

This organization has no public members. You must be a member to see who’s a part of this organization.

Most used topics

Loading…