MtpWeights

Struct MtpWeights 

Source
pub struct MtpWeights {
Show 17 fields pub pre_fc_norm_embedding: DenseWeight, pub pre_fc_norm_hidden: DenseWeight, pub fc: DenseWeight, pub input_layernorm: DenseWeight, pub q_proj: DenseWeight, pub k_proj: DenseWeight, pub v_proj: DenseWeight, pub o_proj: DenseWeight, pub q_norm: DenseWeight, pub k_norm: DenseWeight, pub post_attn_layernorm: DenseWeight, pub moe_gate: DenseWeight, pub shared_expert: DenseExpertWeight, pub shared_expert_gate: DenseWeight, pub experts: Vec<DenseExpertWeight>, pub dense_ffn: Option<DenseExpertWeight>, pub norm: DenseWeight,
}
Expand description

MTP (Multi-Token Prediction) head weights (all BF16 from safetensors).

Single decoder layer + concat projection. All projection weights are BF16 and get quantized to NVFP4 at load time by the weight loader.

Fields§

§pre_fc_norm_embedding: DenseWeight

RMSNorm on token embedding before concat: [hidden_size] BF16.

§pre_fc_norm_hidden: DenseWeight

RMSNorm on target hidden state before concat: [hidden_size] BF16.

§fc: DenseWeight

Concat projection: [hidden_size, 2*hidden_size] BF16.

§input_layernorm: DenseWeight

Input layernorm for the attention layer: [hidden_size] BF16.

§q_proj: DenseWeight

Attention projections (all BF16).

§k_proj: DenseWeight§v_proj: DenseWeight§o_proj: DenseWeight§q_norm: DenseWeight§k_norm: DenseWeight§post_attn_layernorm: DenseWeight

Post-attention layernorm: [hidden_size] BF16.

§moe_gate: DenseWeight

MoE router gate: [num_experts, hidden_size] BF16. NULL when dense_ffn is Some (dense FFN MTP head).

§shared_expert: DenseExpertWeight

Shared expert (BF16). NULL fields when dense_ffn is Some.

§shared_expert_gate: DenseWeight

Shared expert gate: [1, hidden_size] BF16. NULL when dense_ffn is Some.

§experts: Vec<DenseExpertWeight>

Per-expert weights (512 experts, BF16). Empty when dense_ffn is Some.

§dense_ffn: Option<DenseExpertWeight>

Dense FFN triple (gate_proj, up_proj, down_proj) — used by MTP heads bundled with dense (non-MoE) FP8 checkpoints, e.g. Qwen/Qwen3.6-27B-FP8. When Some, the MoE fields above are unused and the forward path takes the dense MLP shortcut.

§norm: DenseWeight

Final output RMSNorm: [hidden_size] BF16.

Auto Trait Implementations§

Blanket Implementations§

Source§

impl<T> Any for T
where T: 'static + ?Sized,

Source§

fn type_id(&self) -> TypeId

Gets the TypeId of self. Read more
Source§

impl<T> Borrow<T> for T
where T: ?Sized,

Source§

fn borrow(&self) -> &T

Immutably borrows from an owned value. Read more
Source§

impl<T> BorrowMut<T> for T
where T: ?Sized,

Source§

fn borrow_mut(&mut self) -> &mut T

Mutably borrows from an owned value. Read more
Source§

impl<T> From<T> for T

Source§

fn from(t: T) -> T

Returns the argument unchanged.

§

impl<T> Instrument for T

§

fn instrument(self, span: Span) -> Instrumented<Self>

Instruments this type with the provided [Span], returning an Instrumented wrapper. Read more
§

fn in_current_span(self) -> Instrumented<Self>

Instruments this type with the current Span, returning an Instrumented wrapper. Read more
Source§

impl<T, U> Into<U> for T
where U: From<T>,

Source§

fn into(self) -> U

Calls U::from(self).

That is, this conversion is whatever the implementation of From<T> for U chooses to do.

Source§

impl<T, U> TryFrom<U> for T
where U: Into<T>,

Source§

type Error = Infallible

The type returned in the event of a conversion error.
Source§

fn try_from(value: U) -> Result<T, <T as TryFrom<U>>::Error>

Performs the conversion.
Source§

impl<T, U> TryInto<U> for T
where U: TryFrom<T>,

Source§

type Error = <U as TryFrom<T>>::Error

The type returned in the event of a conversion error.
Source§

fn try_into(self) -> Result<U, <U as TryFrom<T>>::Error>

Performs the conversion.
§

impl<V, T> VZip<V> for T
where V: MultiLane<T>,

§

fn vzip(self) -> V

§

impl<T> WithSubscriber for T

§

fn with_subscriber<S>(self, subscriber: S) -> WithDispatch<Self>
where S: Into<Dispatch>,

Attaches the provided Subscriber to this type, returning a [WithDispatch] wrapper. Read more
§

fn with_current_subscriber(self) -> WithDispatch<Self>

Attaches the current default Subscriber to this type, returning a [WithDispatch] wrapper. Read more