Files
Sylpheed/tools/ppc-manual/vmx/vmsumshs.md
sim 21161154d1
All checks were successful
CI / Native — linux (pull_request) Successful in 2h2m48s
CI / WASM — Web (pull_request) Successful in 29m51s
CI / Formatting (pull_request) Successful in 1m36s
fix(ppc-manual): 543 dead links, from two generator bugs and wrong relative paths
- Category pages linked each family as `<slug>.md`, relative to categories/,
  where no family page lives. They now link `../<category>/<slug>.md`.
- Form pages linked a member into its *own* category directory, so every
  VMX128 sibling (`vsldoi128`) pointed at vmx128/ although its family page is
  under vmx/. They now link into the family's directory.
- Hand-written "Related" and sibling mentions linked other categories' pages
  as if they were in the same directory. 109 are retargeted through the page
  index; 29 that pointed a family page at itself (`vrefp128` on vrefp.md) and
  6 naming instructions the manual has no page for are plain text now.

Regenerated at the existing Canary pin (f21ebd49e): upstream has moved on, and
re-pinning belongs in its own change. The generator reports 0 family pages
changed and is idempotent; the only dead links left are TEMPLATE.md's
placeholders.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
2026-09-16 22:37:12 +02:00

5.5 KiB
Raw Blame History

vmsumshs — Vector Multiply-Sum Signed Half Word Saturate

Category: VMX (Altivec) · Form: VA · Opcode: 0x10000029

Assembler Mnemonics

Mnemonic XML entry Flags Description
vmsumshs vmsumshs Vector Multiply-Sum Signed Half Word Saturate

Syntax

vmsumshs [VD], [VA], [VB], [VC]

Encoding

vmsumshs — form VA

  • Opcode word: 0x10000029
  • Primary opcode (bits 05): 4
  • Extended opcode: 41
  • Synchronising: no
Bits Field Meaning
05 OPCD primary opcode (4)
610 VRT destination vector register
1115 VRA source A
1620 VRB source B
2125 VRC source C / shift
2631 XO extended opcode (6 bits)

Operands

Field Role Description
VA vmsumshs: read Source A vector register.
VB vmsumshs: read Source B vector register.
VC vmsumshs: read Source C vector register / 3-bit selector.
VD vmsumshs: write Destination vector register.
VSCR vmsumshs: write Vector Status and Control Register (NJ/SAT bits).

Register Effects

vmsumshs

  • Reads (always): VA, VB, VC
  • Reads (conditional): none
  • Writes (always): VD, VSCR
  • Writes (conditional): none

Status-Register Effects

  • vmsumshs: VSCR[SAT] may be stickied on saturating vector operations.

Operation (pseudocode)

; No hand-written pseudocode for this instruction yet.
; The authoritative semantics are the Canary emitter snapshot under
; Implementation References; about half of Canary's emitters open
; with the PPC-style definition as a comment (`RD <- (RA) + (RB)`).
; Every side effect is also enumerated in the Register Effects and
; Status-Register Effects tables above.

C Translation Example

/* No hand-written C yet. Translate the Canary emitter snapshot   */
/* under Implementation References; its HIR maps directly:        */
/*   f.LoadGPR(n) / f.StoreGPR(n, v)  -> r[n] / r[n] = v          */
/*   f.LoadFPR / StoreFPR, f.LoadVR / StoreVR -> f[n], v[n]        */
/*   f.Load(ea, T), f.Store(ea, v) -> raw read / write; emitters   */
/*     wrap them in f.ByteSwap for the big-endian guest value      */
/*   f.UpdateCR(n, v)  -> CR field n from v's LOW 32 BITS vs 0     */
/*   f.LoadCA / f.StoreCA -> xer.CA;  f.StoreSAT -> vscr.SAT       */
/*   i.XO.RA, i.D.DS, ... -> the bit-fields listed under Operands  */
/* The Register Effects and Status-Register Effects tables above  */
/* enumerate every side effect a faithful translation must emit.  */

Implementation References

vmsumshs

Canary emitter (frozen snapshot @ f21ebd49e9)
int InstrEmit_vmsumshs(PPCHIRBuilder& f, const InstrData& i) {
  XEINSTRNOTIMPLEMENTED();
  return 1;
}

Special Cases & Edge Conditions

  • Signed half-word multiply-sum, saturating. Per word lane:
    VD[i] = clamp(VC[i] + int16(VA[2*i]) * int16(VB[2*i])
                         + int16(VA[2*i+1]) * int16(VB[2*i+1]), INT32_MIN, INT32_MAX)
    
    Two signed-half × signed-half products plus a signed-word accumulator, clamped to int32.
  • Wide-then-clamp ordering. The IBM specification accumulates the full sum and clamps only the final result to int32, which avoids spurious mid-sum saturation that would happen if the products were clamped individually. ⚠️ Canary does not implement vmsumshs: its emitter is XEINSTRNOTIMPLEMENTED, so translating one logs "Unimplemented instr" and, with the default break_on_unimplemented_instructions, breaks.
  • VSCR[SAT] is sticky-set if any of the four lane sums saturates. Cleared only via mtvscr.
  • Big-endian half lanes. Lane 0 is the most-significant half.
  • No XER, no exceptions.
  • Aliasing legal.
  • No VMX128 sibling.
  • Common usage. High-precision dot products, audio FIR taps with overflow detection, signed-pixel filter convolution.
  • vmsumshm — same shape, modulo (no clamp, no SAT flag).
  • vmsumuhs — unsigned half multiply-sum, saturating.
  • vmsummbm, vmsumubm — multiply-sum at byte width.
  • vaddsws — saturating word add for further accumulation.
  • mtvscr / mfvscr — read or clear VSCR[SAT].

IBM Reference