Skip to content
@cxcscmu

cxcscmu

Popular repositories Loading

  1. Craw4LLM Craw4LLM Public

    Official repository for "Craw4LLM: Efficient Web Crawling for LLM Pretraining"

    Python 612 56

  2. RAGViz RAGViz Public

    Official repository for RAGViz: Diagnose and Visualize Retrieval-Augmented Generation [EMNLP 2024]

    TypeScript 82 12

  3. MATES MATES Public

    Official repository for MATES: Model-Aware Data Selection for Efficient Pretraining with Data Influence Models [NeurIPS 2024]

    Python 64 8

  4. Montessori-Instruct Montessori-Instruct Public

    Official repository for Montessori-Instruct: Generate Influential Training Data Tailored for Student Learning [ICLR 2025]

    Python 44 4

  5. ED-Copilot ED-Copilot Public

    Python 8 1

  6. FactMM-RAG FactMM-RAG Public

    Official repository for FactMM-RAG: Fact-Aware Multimodal Retrieval Augmentation for Accurate Medical Radiology Report Generation [NAACL 2025]

    Python 6

Repositories

Showing 10 of 12 repositories
  • FactMM-RAG Public

    Official repository for FactMM-RAG: Fact-Aware Multimodal Retrieval Augmentation for Accurate Medical Radiology Report Generation [NAACL 2025]

    Python 6 MIT 0 0 1 Updated Apr 21, 2025
  • ResearchArena Public
    Python 1 0 0 0 Updated Apr 2, 2025
  • embedding-scope Public

    Interpret and control dense embedding via sparse autoencoder.

    Python 5 MIT 0 0 0 Updated Mar 5, 2025
  • Craw4LLM Public

    Official repository for "Craw4LLM: Efficient Web Crawling for LLM Pretraining"

    Python 612 MIT 56 3 0 Updated Feb 24, 2025
  • CLUE-LLM-site Public
    TypeScript 0 0 0 0 Updated Feb 11, 2025
  • Montessori-Instruct Public

    Official repository for Montessori-Instruct: Generate Influential Training Data Tailored for Student Learning [ICLR 2025]

    Python 44 MIT 4 0 0 Updated Jan 24, 2025
  • RAGViz Public

    Official repository for RAGViz: Diagnose and Visualize Retrieval-Augmented Generation [EMNLP 2024]

    TypeScript 82 MIT 12 1 0 Updated Jan 18, 2025
  • MATES Public

    Official repository for MATES: Model-Aware Data Selection for Efficient Pretraining with Data Influence Models [NeurIPS 2024]

    Python 64 MIT 8 4 0 Updated Nov 14, 2024
  • esae Public
    Python 0 0 0 0 Updated Oct 29, 2024
  • Python 1 0 0 0 Updated Oct 23, 2024

Top languages

Loading…

Most used topics

Loading…