huggingface/diffusers

[Proposals Welcome] Fal Flashpack integration for faster model loading

開放

#12,564 建立於 2025年10月31日

 (4 則留言) (3 個反應) (0 位負責人)Python (4,562 個分叉)batch import
contributions-welcomehelp wantedstale

倉庫指標

星標
 (22,190 顆星)
PR 合併指標
 (平均合併 14天 10小時) (30 天內合併 101 個 PR)

描述

Hey! 👋

We've had a request to explore integrating Fal's Flashpack for faster DiT and Text Encoder loading (https://github.com/huggingface/diffusers/issues/12550). Before we jump into implementation, we wanted to open this up to the community to gather ideas and hear from anyone who's experimented with this.

We'd love your input on:

  1. Performance: Has anyone tried it? What kind of speedups did you see? Are there any performance trade-offs?
  2. Integration Design: How would you approach it if you were to integrating this into Diffusers? Describe your design at a high level - how would we support this in our existing framework and what would the API look like?

We're looking for proposals and ideas rather than PRs at this stage. We're genuinely interested in hearing different approaches and perspectives from the community on this.

Feel free to share your thoughts!

貢獻者指南