M. Zaharia
发表论文 246 篇 · 总被引 72453 次 · h-index 78
代表论文
- FrugalGPT: How to Use Large Language Models While Reducing Cost and Improving Performance (2023 · Trans. Mach. Learn. Res. · 被引 831)
- Demonstrate-Search-Predict: Composing retrieval and language models for knowledge-intensive NLP (2022 · arXiv.org · 被引 397)
- Exploiting Programmatic Behavior of LLMs: Dual-Use Through Standard Security Attacks (2023 · 2024 IEEE Security and Privacy Workshops (SPW) · 被引 392)
- RAFT: Adapting Language Model to Domain Specific RAG (2024 · arXiv.org · 被引 377)
- MegaBlocks: Efficient Sparse Training with Mixture-of-Experts (2022 · Conference on Machine Learning and Systems · 被引 243)
- MoE-Lightning: High-Throughput MoE Inference on Memory-constrained GPUs (2024 · International Conference on Architectural Support for Programming Languages and Operating Systems · 被引 85)