The memory map gave CosmOS 0x0000 through 0x1FFF of Program Memory and applications 0x2000 and above. CosmOS is 8141 bytes at the previous commit, which is fifty one bytes short of the line, and the next thing added to it went over. GOING OVER DOES NOT FAIL WHERE IT HAPPENS. Nothing enforces the division: an application says where it goes with #Base and the loader puts it there, so a CosmOS that has grown past 0x1FFF simply has the next program loaded written over the end of it. What breaks is whichever part of the shell that program happened to cover, at whatever later moment somebody uses it. It turned up here as the monitor's assemble command answering "I do not know" to valid instructions, several commands into a session, on a machine that had booted perfectly well. Both halves are doubled: applications now start at 0x4000 in Program Memory and 0x2000 in Data Memory. That is 16K of code and 8K of data for the system, against the 8775 and 2948 it uses today. Both were on the same trajectory, and moving them together means the twenty files that say #Base are edited once rather than twice. The standalone loader's loadable.asm keeps its old base: it belongs to the loader CosmOS grew out of, not to CosmOS, and its addresses answer to a different program. The unbased-segment diagnostic keeps its old base too - it exists to produce an error message that names the address, and the message is what is recorded. Tests/docs.sh now reads the two limits out of the table in the README and measures both segments against them. It reads them rather than being told them because the table is the specification, and this is the second time in this project that the thing nobody checked is the thing that rotted. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com> Claude-Session: https://claude.ai/code/session_01E2JrLzFvuFX9fgi1LDRjrW
178 lines
3.1 KiB
NASM
178 lines
3.1 KiB
NASM
; The assembler's front end on its own, one line per token.
|
|
;
|
|
; Prints the line each token started on, what the token turned out to be, how many bytes
|
|
; it will come to, and the token itself between brackets so that whitespace at either end
|
|
; would show if any ever leaked in.
|
|
;
|
|
; WHAT A TOKEN IS is the part worth checking here rather than at the far end. A wrong
|
|
; classification does not produce a wrong byte in an obvious place - it produces a right
|
|
; looking program of the wrong length, with everything after it shifted, and by then the
|
|
; only symptom is that a label points at the middle of an instruction.
|
|
;
|
|
; Written by Anachronaut
|
|
|
|
#Include services.asm
|
|
|
|
#Program
|
|
|
|
#Base 0x4000
|
|
|
|
start:
|
|
SETD.0 Wanted
|
|
INIB 0d23
|
|
SWI osArgument
|
|
SETD.0 Wanted
|
|
LDA.0
|
|
BRA nothingAsked
|
|
|
|
SETD.0 Wanted
|
|
CALL srcOpen
|
|
BNQ noFile
|
|
|
|
tokenLoop:
|
|
CALL tokNext
|
|
BNQ tokensDone
|
|
|
|
SETD.0 TokLine
|
|
LDA.0
|
|
INCD.0
|
|
LDB.0
|
|
SWI osPrintNumber
|
|
|
|
CALL clsToken
|
|
BNQ badToken
|
|
|
|
; Which of the six it turned out to be. The names are four characters and a space, so
|
|
; the columns line up without any counting.
|
|
SETD.0 TypeNames
|
|
SETD.2 ClsType
|
|
LDA.2
|
|
CALL nameOfType
|
|
SETD.1 NamePointer
|
|
LDD.0.1
|
|
SWI osPrintString
|
|
|
|
SETD.0 ClsLength
|
|
LDA.0
|
|
INCD.0
|
|
LDB.0
|
|
SWI osPrintNumber ; Sixteen bits: a string of 255 characters is 256 bytes long.
|
|
|
|
SETD.0 OpenMark
|
|
SWI osPrintString
|
|
SETD.0 TokText
|
|
SWI osPrintString
|
|
SETD.0 CloseMark
|
|
SWI osPrintString
|
|
BRI tokenLoop
|
|
|
|
badToken:
|
|
SETD.0 StoppedText
|
|
SWI osPrintString
|
|
SWI osExit
|
|
|
|
; The name of type A, out of a table of fixed width entries so that no pointer arithmetic
|
|
; is needed beyond a multiply by the width.
|
|
nameOfType:
|
|
SETD.0 TypeWidth
|
|
LDB.0
|
|
RSTA
|
|
SETD.0 NameLeft
|
|
STA.0
|
|
SETD.2 ClsType
|
|
LDA.2
|
|
SETD.0 NameOffset
|
|
CALL numZero
|
|
nameLoop:
|
|
SETD.2 ClsType
|
|
LDA.2
|
|
SETD.0 NameLeft
|
|
LDB.0
|
|
CCF
|
|
SUB
|
|
BRQ nameFound
|
|
SETD.0 TypeWidth
|
|
LDA.0
|
|
SETD.0 NameOffset
|
|
CALL numAddByte
|
|
SETD.0 NameLeft
|
|
LDA.0
|
|
INCA
|
|
STA.0
|
|
BRI nameLoop
|
|
nameFound:
|
|
; The answer is left in NamePointer rather than in DP0, because a RET puts Data Pointers
|
|
; 0 to 2 back as they were: a routine cannot hand back a pointer, only write one down.
|
|
SETD.0 TypeNames
|
|
SETD.1 NamePointer
|
|
STD.0.1
|
|
SETD.0 NamePointer
|
|
SETD.2 NameOffset
|
|
CALL numAdd
|
|
RET
|
|
|
|
tokensDone:
|
|
SETD.0 DoneText
|
|
SWI osPrintString
|
|
SWI osExit
|
|
|
|
nothingAsked:
|
|
SETD.0 AskText
|
|
SWI osPrintString
|
|
SWI osExit
|
|
|
|
noFile:
|
|
SETD.0 NoFileText
|
|
SWI osPrintString
|
|
SWI osExit
|
|
|
|
#Data
|
|
|
|
#Base 0x2000
|
|
|
|
Wanted:
|
|
#Reserve 0d23
|
|
|
|
OpenMark:
|
|
" ["
|
|
CloseMark:
|
|
"]
|
|
"
|
|
; Six names of eleven characters each, counting the zero the assembler puts on the end of
|
|
; every string. Written one to a line so that adding a type is adding a line.
|
|
TypeNames:
|
|
" keyword "
|
|
" instr "
|
|
" value "
|
|
" string "
|
|
" label: "
|
|
" label "
|
|
TypeWidth:
|
|
0d11
|
|
NameLeft:
|
|
0x00
|
|
NameOffset:
|
|
0x00 0x00
|
|
NamePointer:
|
|
0x00 0x00
|
|
|
|
StoppedText:
|
|
"---- stopped: the assembler does not understand that
|
|
"
|
|
DoneText:
|
|
"---- no more tokens
|
|
"
|
|
AskText:
|
|
"say which file
|
|
"
|
|
NoFileText:
|
|
"no such file
|
|
"
|
|
|
|
#Include scratch.asm
|
|
#Include numbers.asm
|
|
#Include source.asm
|
|
#Include token.asm
|
|
#Include classify.asm
|
|
#Include table.asm
|