Skip to content
AETRONConstructor
My neuronetsRegistrySpacesDocsFAQ
Sign in
AETRON

A constructor for AI services on AETRON’s decentralised network of GPU miners.

ProductCreate a neuronetModel catalogueDatasetsSpacesMy payouts
ResourcesDocumentationAPITimings & statusesFAQ
Networkaetron.ioPrivacy & verification
© 2026 AETRONTestnet
Models/qwen3.5-9b-w4a16
Q

Qwen3.5-9B (w4a16)

Alibaba / Qwen · Qwen3.5
Create a neuronet with this model →
LLMApache-2.0publicTrusted9.7B params
chatinstructmultilingualreasoningcodingquantizedint4bnb

Ready-made INT4 W4A16 build (compressed-tensors): 9B dense LLM from the Qwen3.5 generation (02.2026). Strong reasoning, coding and multilingual chat.

Qwen3.5-9B

Dense 9B model of the Qwen3.5 generation (released 02.2026). Outperforms the Qwen2.5 line in reasoning, coding and agents.

Strengths

  • Modern generation: strong reasoning & coding
  • 32K context, multilingual
  • Apache-2.0 — fully commercial

Recommended use

Primary general-purpose assistant model. Runs on a single 24GB GPU at bf16 (weights ~19.3 GB).

Create a neuronet with this model →
Overview
Task
LLM · InferenceV1
Family
Qwen3.5
Tier
Trusted
Status
active
Runtime spec
dtype
Int4
attention
FlashAttn2
compile
ReduceOverhead
seed
0
gpu_arch
Sm89
Runs on
Consumer, Datacenter
Hardware
Min VRAM
13 GB
Recommended
14 GB
Tensor parallel
≥1
TEE archs
Sm89, Sm90
Verification defaults
Tier
B
Trust
Level1
Security
Standard
Privacy
P0
Min pulse score
100
License
spdx
Apache-2.0 ↗
Commercial use
✓
Redistribution
✓
Source & provenance
Repo
RedHatAI/Qwen3.5-9B-quantized.w4a16 ↗
Revision
a398088c42
Added
2026-08-03
Verified
2026-08-03