Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

litert-community
/
LFM2.5-Encoder-350M-Prompt-Router

Text Classification
LiteRT
LiteRT
on-device
edge
encoder
routing
zero-shot
liquid
lfm2
lfm2.5
Model card Files Files and versions
xet
Community

Instructions to use litert-community/LFM2.5-Encoder-350M-Prompt-Router with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

  • Libraries
  • LiteRT

    How to use litert-community/LFM2.5-Encoder-350M-Prompt-Router with LiteRT:

    # No code snippets available yet for this library.
    
    # To use this model, check the repository files and the library's documentation.
    
    # Want to help? PRs adding snippets are welcome at:
    # https://github.com/huggingface/huggingface.js
  • Notebooks
  • Google Colab
  • Kaggle
LFM2.5-Encoder-350M-Prompt-Router
1.08 GB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 7 commits
mlboydaisuke's picture
mlboydaisuke
Card update: rank-4 re-export GPU results (full mobile delegation; measured on-device)
6d0c68a verified 1 day ago
  • .gitattributes
    1.52 kB
    initial commit 10 days ago
  • LFM2.5-Encoder-350M-Prompt-Router_fp16.tflite
    713 MB
    xet
    Rank-4 repeat_kv re-export: full mobile-GPU delegation (bitwise-identical CPU outputs) 1 day ago
  • LFM2.5-Encoder-350M-Prompt-Router_wi8fc.tflite
    365 MB
    xet
    Rank-4 repeat_kv re-export: full mobile-GPU delegation (bitwise-identical CPU outputs) 1 day ago
  • LICENSE
    10.6 kB
    LFM2.5-Encoder-350M-Prompt-Router LiteRT: int8 (iPhone-verified bit-exact) + fp16, task-level parity verified 10 days ago
  • README.md
    8.57 kB
    Card update: rank-4 re-export GPU results (full mobile delegation; measured on-device) 1 day ago
  • tokenizer.json
    4.73 MB
    LFM2.5-Encoder-350M-Prompt-Router LiteRT: int8 (iPhone-verified bit-exact) + fp16, task-level parity verified 10 days ago
  • tokenizer_config.json
    306 Bytes
    LFM2.5-Encoder-350M-Prompt-Router LiteRT: int8 (iPhone-verified bit-exact) + fp16, task-level parity verified 10 days ago