Neura-Tech-AI's picture
Create NOTICE
94a4e8e verified
Raw
History Blame Contribute Delete
2.48 kB
Neuron-6x4B-Instruct
Copyright (c) 2026 Neura Tech AI
Neuron-6x4B-Instruct is an open-source Mixture of Experts (MoE) large language
model developed by Neura Tech AI.
This model is built upon the following base models:
- Neura-Tech-AI/Neuron-4B-Instruct
Copyright (c) 2026 Neura Tech AI.
- Qwen/Qwen3-4B-Instruct-2507
Copyright (c) 2024 Alibaba Cloud and the Qwen Team.
- Qwen/Qwen3-4B-Thinking-2507
Copyright (c) 2024 Alibaba Cloud and the Qwen Team.
Neuron-4B-Instruct is based on the Qwen3 architecture and includes additional
modifications and improvements developed by Neura Tech AI.
The original Qwen3 models are licensed under the Apache License,
Version 2.0. A copy of the Apache License is included in the LICENSE file.
Neuron-6x4B-Instruct introduces additional modifications made by
Neura Tech AI, including but not limited to:
- Mixture of Experts (MoE) architecture
- Six-expert sparse routing design
- Dynamic expert routing configuration
- Expert composition and parameter merging
- Instruction tuning improvements
- Identity customization
- Chat template customization
- Alignment improvements
- Reasoning enhancements
- Multilingual capability improvements
- Coding and software engineering optimization
- Tool-calling optimization
- Long-context support optimization
- Dataset improvements
- Branding and documentation
- Model packaging and distribution
These modifications are Copyright (c) 2026
Neura Tech AI.
Developer
Neura Tech AI
Official AI Research & Development Organization
Project
Neuron
Model
Neuron-6x4B-Instruct
Architecture
Sparse Transformer Decoder
Mixture of Experts (MoE)
Qwen3 Architecture
Base Models
- Neura-Tech-AI/Neuron-4B-Instruct
- Qwen/Qwen3-4B-Instruct-2507
- Qwen/Qwen3-4B-Thinking-2507
Model Highlights
- Approximately 24 Billion Total Parameters
- Six Specialized Experts
- Dynamic Sparse Expert Routing
- Instruction-Tuned
- Multilingual Language Model
- Optimized for Reasoning, Coding, Mathematics, Tool Calling,
Agentic Workflows, and Long-Context Understanding
Acknowledgment
We sincerely thank Alibaba Cloud and the Qwen Team for releasing the
Qwen3 model family under the Apache License, Version 2.0. Their work
served as the architectural foundation that enabled the development of
Neuron-4B-Instruct and, subsequently, Neuron-6x4B-Instruct.
This NOTICE file is provided solely for attribution purposes and does
not modify, replace, or supersede the terms of the Apache License,
Version 2.0.