How to use from
vLLM
Install from pip and serve model
# Install vLLM from pip:
pip install vllm
# Start the vLLM server:
vllm serve "wvvss/GPUmatLLM"
# Call the server using curl (OpenAI-compatible API):
curl -X POST "http://localhost:8000/v1/chat/completions" \
	-H "Content-Type: application/json" \
	--data '{
		"model": "wvvss/GPUmatLLM",
		"messages": [
			{
				"role": "user",
				"content": "What is the capital of France?"
			}
		]
	}'
Use Docker
docker model run hf.co/wvvss/GPUmatLLM
Quick Links

GPUmatLLM

A domain-specific large language model for GPU materials science, obtained by LoRA fine-tuning of Qwen2.5-7B-Instruct.

Model details

  • Base model: Qwen2.5-7B-Instruct
  • Adaptation method: LoRA (rank 8)
  • Domain: GPU packaging materials, thermal management, semiconductor substrates, interconnect materials
  • Language: Chinese

Intended use

Research use for domain question answering in GPU materials science.

Limitations

Outputs may contain factual errors and should be verified against primary sources before use in engineering decisions.

Citation

Paper under review. Citation information will be added upon publication.

Downloads last month
103
Safetensors
Model size
8B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for wvvss/GPUmatLLM

Base model

Qwen/Qwen2.5-7B
Adapter
(2600)
this model