Does StereoMath (2024) cite AsTeR/Emacspeak for the pitch-encodes-vertical-position mapping, or present it as novel? (rediscovery vs. acknowledged lineage)
This capture answers the open question at
question-does-stereomath-2024-cite-aster-emacspeak-pitch-mapping-prior-art by
reading StereoMath's full text and reference list directly (arXiv 2501.01404, PDF
extracted and read in full — 5 pages, all sections and all 39 references). The
paper's own tls provenance on fetch was verified, no elevated-suspicion handling
needed, and no recognition signals from the safety spec fired anywhere in the text —
see Safety flags below (none).
Claim: StereoMath cites Raman's Emacspeak — not AsTeR — and only as generic sonification prior art, not for the pitch mapping specifically
In its Related Work section, StereoMath writes: "Luckily, work in other domains has successfully overcome similar challenges, by using sonification [9, 18] (e.g., in SAS Graphics Accelerator [6] and EmacsSpeak [29]), as well as employing modality switching [28], which has been shown to reduce cognitive load [11, 12, 20]." Reference [29] resolves to: "T. V. Raman. Emacspeak—a speech interface. In Proceedings of the SIGCHI conference on Human factors in computing systems common ground - CHI '96... ACM Press, 66–71." This is the paper's only citation of Raman, and it appears in a single clause listing sonification as one successful strategy from "other domains" generally — it is not connected to, or invoked near, any discussion of pitch or vertical position. (Tier 1 — primary text, direct quote.)
Claim: StereoMath's own description of the pitch-to-vertical-position mechanism carries no citation and explicitly frames it as novel
In the implementation section (4.1, "Spatial Sound," addressing design goal DG1), the paper states: "Arguably some of our most important innovations come from our spatial sound and spatial navigation. Here, every time the user interacts in a way that produces sound (e.g., navigation or typing), we modify the audio to match the spatial left-right location in the equation's final rendering. We also change the pitch to match the vertical up/down location. In our co-design sessions, this improved ease of comprehension, and provided a novel experience that was previously not commonly found elsewhere." No citation is attached to this passage, either at first mention (design goal DG1, §3.2) or here at implementation. This is the paper's own novelty framing for the exact mapping that AsTeR implemented in 1994. (Tier 1 — primary text, direct quote.)
Claim: AsTeR itself is never named anywhere in StereoMath's paper — only Emacspeak, a related but distinct Raman project, appears
A full read of the text and all 39 references confirms the string "AsTeR" does not occur anywhere in the paper. Raman appears exactly once, as author of reference [29], Emacspeak (CHI '96) — a general-purpose audio desktop / speech interface for Emacs, not the math-specific pitched-audio reading system (AsTeR) that is the direct ancestor of the pitch-for-vertical mapping. The question's framing ("AsTeR/Emacspeak") treats the two as interchangeable prior art; StereoMath's own citation practice does not reach back to AsTeR specifically at all. (Tier 1 — primary text, absence confirmed by direct full-text read of the extracted PDF.)
Claim: the question resolves as partial, generic lineage acknowledgment plus an uncited, explicit novelty claim for the specific mechanism — not a clean case of either rediscovery or acknowledged lineage
Put together, the three findings above show the paper doing both things the open question posed as alternatives, in different places: it credits the sonification genre broadly (citing Emacspeak, not AsTeR, as one example of prior success "in other domains"), while separately and explicitly claiming the specific pitch-encodes-vertical-position mechanism as one of its "most important innovations" and "a novel experience... not commonly found elsewhere," with no citation at that point and no citation of AsTeR anywhere in the paper. This directly answers observation-math-sonification-pitch-for-vertical-recurred-1994-2024's open question: the paper's hop-chain tension ("cites Emacspeak as prior art" vs. "treats the mapping as an innovation") is not a contradiction to resolve but an accurate description of the paper — both are literally true, of different passages, about different sources (Emacspeak vs. AsTeR). The specific pitch-for-vertical mechanism is best read as an independent rediscovery: it is not attributed to AsTeR, and the one citation to Raman's other project does not carry the mechanism with it. (Tier 1 — synthesis of the three primary-text findings above, not a new external fact.)
Further leads
- Whether Emacspeak (1996, the system StereoMath does cite) itself implements any pitch-for-structure mapping, distinct from AsTeR's math-specific mechanism — would clarify whether StereoMath's citation is even adjacent to the right idea, or purely coincidental. Source: StereoMath ref [29], CHI '96 paper (not yet read).
- StereoMath's other cited sonification prior art — Barrass & Kramer "Using Sonification" [9] and Holloway et al. "Infosonics" [18] — neither yet checked for pitch-vertical mappings of their own. Source: arXiv 2501.01404 reference list.
- The arXiv record for 2501.01404 is versioned v2, dated 15 Mar 2026, revising the original ASSETS '24 (Oct 2024) presentation; not checked whether citations changed between versions. Source: arXiv 2501.01404 header line "arXiv:2501.01404v2 [cs.HC] 15 Mar 2026."
Entity candidates
- Kenneth Ge — person — StereoMath co-author (sighted researcher/developer); this was his first publication per the paper's acknowledgments.
- JooYoung Seo — person — StereoMath co-author, School of Information Sciences, University of Illinois Urbana-Champaign; accessibility researcher.
- Emacspeak — concept/term — Raman's 1996 CHI speech-interface/audio-desktop project, the only Raman work StereoMath actually cites; distinct from AsTeR and worth its own note on what it does and doesn't share with AsTeR's math-specific mechanism.
- ASSETS (ACM SIGACCESS Conference on Computers and Accessibility) — concept/term — the venue for both this paper and other cited accessibility work; recurs across the vault's blind-mathematicians cluster.
Safety flags
None. The only source fetched this session was the arXiv PDF (2501.01404), retrieved via extract_pdf with tls: "verified". Full text was read; no addressed-to-AI language, override language, claimed authority, tier self-assignment, file-system instructions, credential requests, or urgency framing appeared anywhere in the paper.