amd/DeepSeek-V4-Flash-NVFP4开源了新模型
根据官方来源,amd/DeepSeek-V4-Flash-NVFP4开源了新模型。详细信息请以原始来源为准。
证据来源
查看来源摘录
AMD amd/DeepSeek-V4-Flash-NVFP4 amd/DeepSeek-V4-Flash-NVFP4 https://huggingface.co/amd/DeepSeek-V4-Flash-NVFP4
查看来源摘录
amd/DeepSeek-V4-Flash-NVFP4 · Hugging Face Hugging Face (https://huggingface.co/) Models (https://huggingface.co/models) Datasets (https://huggingface.co/datasets) Spaces (https://huggingface.co/spaces) Buckets new (https://huggingface.co/storage) Docs (https://huggingface.co/docs) Enterprise (https://huggingface.co/enterprise) Pricing (https://huggingface.co/pricing) Website Tasks (https://huggingface.co/tasks) HuggingChat (https://huggingface.co/chat) Collections (https://huggingface.co/collections) Languages (https://huggingface.co/languages) Organizations (https://huggingface.co/organizations) Community Blog (https://huggingface.co/blog) Posts (https://huggingface.co/posts) Daily Papers (https://huggingface.co/papers) Hardware (https://huggingface.co/hardware) Learn (https://huggingface.co/learn) Discord (https://huggingface.co/join/discord) Forum (https://discuss.huggingface.co/) GitHub (https://github.com/huggingface) Solutions Team & Enterprise (https://huggingface.co/enterprise) Hugging Face PRO (https://huggingface.co/pro) Enterprise Support (https://huggingface.co/support) Inference Providers (https://huggingface.co/inference/models) Inference Endpoints (https://huggingface.co/inference-endpoints) Storage Buckets (https://huggingface.co/storage) Log In (https://huggingface.co/login) Sign Up (https://huggingface.co/join) (https://huggingface.co/amd) amd (https://huggingface.co/amd) / DeepSeek-V4-Flash-NVFP4 (https://huggingface.co/amd/DeepSeek-V4-Flash-NVFP4) like 2 Follow AMD 3.09k Text Generation (https://huggingface.co/models?pipeline_tag=text-generation) Transformers (https://huggingface.co/models?library=transformers) Safetensors (https://huggingface.co/models?library=safetensors) English (https://huggingface.co/models?language=en) Chinese (https://huggingface.co/models?language=zh) deepseek_v4 (https://huggingface.co/models?other=deepseek_v4) nvfp4 (https://huggingface.co/models?other=nvfp4) ocp-mx (https://huggingface.co/models?other=ocp-mx) quantization (https://huggingface.co/models?other=quantization) amd-quark (https://huggingface.co/models?other=amd-quark) Mixture of Experts (https://huggingface.co/models?other=moe) deepseek (https://huggingface.co/models?other=deepseek) 8-bit precision (https://huggingface.co/models?other=8-bit) quark (https://huggingface.co/models?other=quark) License: mit Model card (https://huggingface.co/amd/DeepSeek-V4-Flash-NVFP4) Files Files and versions xet (https://huggingface.co/amd/DeepSeek-V4-Flash-NVFP4/tree/main) Community 1 (https://huggingface.co/amd/DeepSeek-V4-Flash-NVFP4/discussions) Deploy Copy to bucket new Use this model Instructions to use amd/DeepSeek-V4-Flash-NVFP4 with libraries, inference providers, notebooks, and local apps. Follow these links to get started. Libraries Transformers (https://huggingface.co/amd/DeepSeek-V4-Flash-NVFP4?library=transformers) How to use amd/DeepSeek-V4-Flash-NVFP4 with Transformers: # Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("text-generation", model="amd/DeepSeek-V4-Flash-NVFP4") # Load model directly from transformers import AutoTokenizer, AutoModelForCausalLM tokenizer = AutoTokenizer.from_pretrained("amd/DeepSeek-V4-Flash-NVFP4") model = AutoModelForCausalLM.from_pretrained("amd/DeepSeek-V4-Flash-NVFP4", device_map="auto") Notebooks Google Colab (https://huggingface.co/amd/DeepSeek-V4-Flash-NVFP4/colab) Kaggle (https://huggingface.co/amd/DeepSeek-V4-Flash-NVFP4/kaggle) Local Apps Settings (https://huggingface.co/settings/local-apps) vLLM (https://huggingface.co/amd/DeepSeek-V4-Flash-NVFP4?local-app=vllm) How to use amd/DeepSeek-V4-Flash-NVFP4 with vLLM: Install from pip and serve model # Install vLLM from pip: pip install vllm # Start the vLLM server: vllm serve "amd/DeepSeek-V4-Flash-NVFP4" # Call the server using curl (OpenAI-compatible API): curl -X POST "http://localhost:8000/v1/completions" \ -H "Content-Type:
来自 星盘大模型百科