Skip to content
@cxcscmu

cxcscmu

Popular repositories Loading

  1. Craw4LLM Craw4LLM Public

    Official repository for "Craw4LLM: Efficient Web Crawling for LLM Pretraining"

    Python 473 41

  2. RAGViz RAGViz Public

    Official repository for RAGViz: Diagnose and Visualize Retrieval-Augmented Generation [EMNLP 2024]

    TypeScript 79 11

  3. MATES MATES Public

    Official repository for MATES: Model-Aware Data Selection for Efficient Pretraining with Data Influence Models [NeurIPS 2024]

    Python 59 7

  4. Montessori-Instruct Montessori-Instruct Public

    Official repository for Montessori-Instruct: Generate Influential Training Data Tailored for Student Learning [ICLR 2025]

    Python 41 3

  5. ED-Copilot ED-Copilot Public

    Python 6 1

  6. LongEmbeddingAnalysis LongEmbeddingAnalysis Public

    Python 3

Repositories

Showing 10 of 10 repositories
  • Craw4LLM Public

    Official repository for "Craw4LLM: Efficient Web Crawling for LLM Pretraining"

    cxcscmu/Craw4LLM’s past year of commit activity
    Python 473 MIT 41 0 0 Updated Feb 24, 2025
  • FactMM-RAG Public

    Official repository for FactMM-RAG: Fact-Aware Multimodal Retrieval Augmentation for Accurate Medical Radiology Report Generation [NAACL 2025]

    cxcscmu/FactMM-RAG’s past year of commit activity
    Python 1 MIT 0 0 1 Updated Feb 22, 2025
  • Montessori-Instruct Public

    Official repository for Montessori-Instruct: Generate Influential Training Data Tailored for Student Learning [ICLR 2025]

    cxcscmu/Montessori-Instruct’s past year of commit activity
    Python 41 MIT 3 0 0 Updated Jan 24, 2025
  • RAGViz Public

    Official repository for RAGViz: Diagnose and Visualize Retrieval-Augmented Generation [EMNLP 2024]

    cxcscmu/RAGViz’s past year of commit activity
    TypeScript 79 MIT 11 1 0 Updated Jan 18, 2025
  • embedding-scope Public

    Interpret and control dense embedding via sparse autoencoder.

    cxcscmu/embedding-scope’s past year of commit activity
    Python 3 MIT 0 0 0 Updated Dec 31, 2024
  • MATES Public

    Official repository for MATES: Model-Aware Data Selection for Efficient Pretraining with Data Influence Models [NeurIPS 2024]

    cxcscmu/MATES’s past year of commit activity
    Python 59 MIT 7 3 0 Updated Nov 14, 2024
  • esae Public
    cxcscmu/esae’s past year of commit activity
    Python 0 0 0 0 Updated Oct 29, 2024
  • cxcscmu/InContextDataAttribution’s past year of commit activity
    Python 1 0 0 0 Updated Oct 23, 2024
  • ED-Copilot Public
    cxcscmu/ED-Copilot’s past year of commit activity
    Python 6 1 0 0 Updated Aug 23, 2024
  • cxcscmu/LongEmbeddingAnalysis’s past year of commit activity
    Python 3 0 0 0 Updated Jun 20, 2024

Most used topics

Loading…