Qwen3-4B-PlumEsper
4
4.0B
2 languages
dataset:sequelbox/Mitakihara-DeepSeek-R1-0528
by
sequelbox
Language Model
OTHER
4B params
New
4 downloads
Early-stage
Edge AI:
Mobile
Laptop
Server
9GB+ RAM
Mobile
Laptop
Server
Quick Summary
This is a merge of pre-trained language models created using mergekit, combining the specialty and general reasoning skills of Esper 3 4b and Shining Valiant 3 4b.
Device Compatibility
Mobile
4-6GB RAM
Laptop
16GB RAM
Server
GPU
Minimum Recommended
4GB+ RAM
Code Examples
Configurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BConfigurationyaml
merge_method: della
dtype: bfloat16
parameters:
normalize: true
models:
- model: ValiantLabs/Qwen3-4B-Esper3
parameters:
density: 0.5
weight: 0.3
- model: ValiantLabs/Qwen3-4B-ShiningValiant3
parameters:
density: 0.5
weight: 0.3
base_model: Qwen/Qwen3-4BDeploy This Model
Production-ready deployment in minutes
Together.ai
Instant API access to this model
Production-ready inference API. Start free, scale to millions.
Try Free APIReplicate
One-click model deployment
Run models in the cloud with simple API. No DevOps required.
Deploy NowDisclosure: We may earn a commission from these partners. This helps keep LLMYourWay free.