Files
Sylpheed/tools/ppc-manual/vmx/vlogefp.md
sim f3c512f2ab docs(ppc-manual): check every xenia-rs claim against Canary's source
The hand-written parts of the manual still described how the retired
xenia-rs interpreter behaved: its snapshots, Rust casts and helpers. Each of
those 490 statements is now either restated as what Canary's emitters and
x64 backend actually do (at the pinned canary_experimental commit), or
dropped where it only made sense for xenia-rs.

Checking them turned up claims that were wrong, not just outdated:

- VSCR[SAT] is never modelled in Canary (DID_SATURATE is a stub and mfvscr
  cannot see it); the pages said saturating ops set it stickily.
- Canary does not implement lswi/lswx/stswi/stswx, dcbi, mtfsb0/mtfsb1,
  vmsum*, vmhaddshs, vupkhpx/vupklpx, and most SPRs; pages described them
  as working.
- Traps evaluate TO in Canary; stvebx/stvehx/stvewx store one element, not
  16 bytes; mtmsrd writes only EE; fres/frsqrte/vrsqrtefp precision claims
  and the stfs "rounds under RN / sets FPSCR" claim contradicted the spec.
- Reservations are a 64 KiB block bitmap plus a value compare, not
  per-address tracking.

Claims that neither Canary's source nor a public spec settles are marked
unverified (NI at boot, vmaddcfp128 operand order, estimate bit-exactness).

Generated regions are untouched; re-running the generator changes nothing.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
2026-09-16 21:52:38 +02:00

7.0 KiB
Raw Permalink Blame History

vlogefp — Vector Log2 Estimate Floating Point

Category: VMX (Altivec) · Form: VX · Opcode: 0x100001ca

Assembler Mnemonics

Mnemonic XML entry Flags Description
vlogefp vlogefp Vector Log2 Estimate Floating Point
vlogefp128 vlogefp128 Vector128 Log2 Estimate Floating Point

Syntax

vlogefp [VD], [VB]
vlogefp128 [VD], [VB]

Encoding

vlogefp — form VX

  • Opcode word: 0x100001ca
  • Primary opcode (bits 05): 4
  • Extended opcode: 458
  • Synchronising: no
Bits Field Meaning
05 OPCD primary opcode (4)
610 VRT/VD destination vector register
1115 VRA/VA source A vector register
1620 VRB/VB source B vector register
2131 XO extended opcode (11 bits)

vlogefp128 — form VX128_3

  • Opcode word: 0x180006f0
  • Primary opcode (bits 05): 6
  • Extended opcode: 1776
  • Synchronising: no
Bits Field Meaning
05 OPCD primary opcode (6)
610 VD128l destination low 5 bits
1115 IMM 5-bit immediate
1620 VB128l source B low 5 bits
2127 XO extended opcode
2829 VD128h destination high 2 bits
3031 VB128h source B high 2 bits

Operands

Field Role Description
VB vlogefp: read; vlogefp128: read Source B vector register.
VD vlogefp: write; vlogefp128: write Destination vector register.

Register Effects

vlogefp

  • Reads (always): VB
  • Reads (conditional): none
  • Writes (always): VD
  • Writes (conditional): none

vlogefp128

  • Reads (always): VB
  • Reads (conditional): none
  • Writes (always): VD
  • Writes (conditional): none

Status-Register Effects

No condition-register or status-register effects.

Operation (pseudocode)

; No hand-written pseudocode for this instruction yet.
; The authoritative semantics are the Canary emitter snapshot under
; Implementation References; about half of Canary's emitters open
; with the PPC-style definition as a comment (`RD <- (RA) + (RB)`).
; Every side effect is also enumerated in the Register Effects and
; Status-Register Effects tables above.

C Translation Example

/* No hand-written C yet. Translate the Canary emitter snapshot   */
/* under Implementation References; its HIR maps directly:        */
/*   f.LoadGPR(n) / f.StoreGPR(n, v)  -> r[n] / r[n] = v          */
/*   f.LoadFPR / StoreFPR, f.LoadVR / StoreVR -> f[n], v[n]        */
/*   f.Load(ea, T), f.Store(ea, v) -> raw read / write; emitters   */
/*     wrap them in f.ByteSwap for the big-endian guest value      */
/*   f.UpdateCR(n, v)  -> CR field n from v's LOW 32 BITS vs 0     */
/*   f.LoadCA / f.StoreCA -> xer.CA;  f.StoreSAT -> vscr.SAT       */
/*   i.XO.RA, i.D.DS, ... -> the bit-fields listed under Operands  */
/* The Register Effects and Status-Register Effects tables above  */
/* enumerate every side effect a faithful translation must emit.  */

Implementation References

vlogefp

Canary emitter (frozen snapshot @ f21ebd49e9)
int InstrEmit_vlogefp(PPCHIRBuilder& f, const InstrData& i) {
  return InstrEmit_vlogefp_(f, i.VX.VD, i.VX.VB);
}

// ── delegates to (src/xenia/cpu/ppc/ppc_emit_altivec.cc:773) ──
int InstrEmit_vlogefp_(PPCHIRBuilder& f, uint32_t vd, uint32_t vb) {
  // (VD) <- log2(VB)
  Value* v = f.Log2(f.LoadVR(vb));
  f.StoreVR(vd, v);
  return 0;
}

vlogefp128

Canary emitter (frozen snapshot @ f21ebd49e9)
int InstrEmit_vlogefp128(PPCHIRBuilder& f, const InstrData& i) {
  return InstrEmit_vlogefp_(f, VX128_3_VD128, VX128_3_VB128);
}

// ── delegates to (src/xenia/cpu/ppc/ppc_emit_altivec.cc:773) ──
int InstrEmit_vlogefp_(PPCHIRBuilder& f, uint32_t vd, uint32_t vb) {
  // (VD) <- log2(VB)
  Value* v = f.Log2(f.LoadVR(vb));
  f.StoreVR(vd, v);
  return 0;
}

Special Cases & Edge Conditions

  • Per-lane base-2 logarithm. Each of the four word lanes computes VD[i] = log2(VB[i]) in binary32. Note: the IBM manual specifies a low-precision estimate (≤ 1/32 ULP relative error). Canary calls the host's std::log2 per lane, which is full-precision; hardware-precise programs may observe small numerical differences.
  • Use vexptefp for the inverse. Pair gives 2^(log2(x)) ≈ x for positive finite x.
  • Big-endian word lanes. Lane 0 is the most-significant word.
  • NaN, negatives, zero, ±∞. log2(negative) and log2(NaN) produce NaN; log2(+0) = -∞; log2(-0) = -∞ (per IEEE-754); log2(+∞) = +∞. None of these stickies VSCR[SAT] — float ops never touch SAT.
  • No exception, no VSCR[SAT] change, no XER change.
  • VMX128 sibling (vlogefp128). Identical semantics with the extended encoding.
  • Natural log via change-of-base. ln(x) = log2(x) * (1 / log2(e)) — multiply by a constant with vmaddfp.
  • vexptefp — base-2 exponent (the inverse).
  • vrefp — reciprocal estimate.
  • vrsqrtefp — reciprocal-square-root estimate.
  • vmaddfp — fused multiply-add for change-of-base scaling.

IBM Reference