Skip to content
@cxcscmu

cxcscmu

Popular repositories Loading

  1. Craw4LLM Craw4LLM Public

    Official repository for "Craw4LLM: Efficient Web Crawling for LLM Pretraining"

    Python 578 52

  2. RAGViz RAGViz Public

    Official repository for RAGViz: Diagnose and Visualize Retrieval-Augmented Generation [EMNLP 2024]

    TypeScript 81 11

  3. MATES MATES Public

    Official repository for MATES: Model-Aware Data Selection for Efficient Pretraining with Data Influence Models [NeurIPS 2024]

    Python 61 7

  4. Montessori-Instruct Montessori-Instruct Public

    Official repository for Montessori-Instruct: Generate Influential Training Data Tailored for Student Learning [ICLR 2025]

    Python 41 3

  5. ED-Copilot ED-Copilot Public

    Python 6 1

  6. LongEmbeddingAnalysis LongEmbeddingAnalysis Public

    Python 3

Repositories

Showing 10 of 10 repositories
  • FactMM-RAG Public

    Official repository for FactMM-RAG: Fact-Aware Multimodal Retrieval Augmentation for Accurate Medical Radiology Report Generation [NAACL 2025]

    Python 2 MIT 0 0 1 Updated Mar 9, 2025
  • embedding-scope Public

    Interpret and control dense embedding via sparse autoencoder.

    Python 3 MIT 0 0 0 Updated Mar 5, 2025
  • Craw4LLM Public

    Official repository for "Craw4LLM: Efficient Web Crawling for LLM Pretraining"

    Python 578 MIT 52 3 0 Updated Feb 24, 2025
  • Montessori-Instruct Public

    Official repository for Montessori-Instruct: Generate Influential Training Data Tailored for Student Learning [ICLR 2025]

    Python 41 MIT 3 0 0 Updated Jan 24, 2025
  • RAGViz Public

    Official repository for RAGViz: Diagnose and Visualize Retrieval-Augmented Generation [EMNLP 2024]

    TypeScript 81 MIT 11 1 0 Updated Jan 18, 2025
  • MATES Public

    Official repository for MATES: Model-Aware Data Selection for Efficient Pretraining with Data Influence Models [NeurIPS 2024]

    Python 61 MIT 7 3 0 Updated Nov 14, 2024
  • esae Public
    Python 0 0 0 0 Updated Oct 29, 2024
  • Python 1 0 0 0 Updated Oct 23, 2024
  • ED-Copilot Public
    Python 6 1 0 0 Updated Aug 23, 2024
  • Python 3 0 0 0 Updated Jun 20, 2024

Top languages

Loading…

Most used topics

Loading…