d0p1/ack - Cute Engineering : Cute solutions to hard problems

d0p1/ack

Author	SHA1	Message	Date
George Koehler	a20b87ca01	In ego, put both words and double-words in reg_float. The size of a reg_float isn't in the descr file, so ego doesn't know. PowerPC and SPARC are the only arches with floating-point registers in their descr files. PowerPC and SPARC registers can hold both 4-byte and 8-byte floats, so I want ego to do both sizes. This might break our SPARC code expander because ego doesn't know that 8-byte values take 2 registers in SPARC. (So ego might allocate too many registers and deallocate too much stack space.) We don't build the SPARC code expander, and its descr file is already wrong: its list of register save costs is too short, so ego will read past the end of the array. This commit doesn't fix the problem with ego and PowerPC ncg. Right now, ncg refuses to put 4-byte floats in registers, but ego expects them to get registers and deallocates their stack space. So ncg emits programs that use the deallocated space, and the values of 4-byte floats become corrupt.	2017-02-16 19:55:52 -05:00
George Koehler	cbe5d8640b	Add floating-point register variables to PowerPC ncg. Use f14 to f31 as register variables for 8-byte double-precison. There are no regvars for 4-byte double precision, because all regvar(reg_float) must have the same size. I expect more programs to prefer 8-byte double precision. Teach mach/powerpc/ncg/mach.c to emit stfd and lfd instructions to save and restore 8-byte regvars. Delay emitting the function prolog until f_regsave(), so we can use one addi to make stack space for both local vars and saved registers. Be more careful with types in mach.c; don't assume that int and long and full are the same. In ncg table, add f14 to f31 as register variables, and some rules to use them. Add rules to put the result of fadd, fsub, fmul, fdiv, fneg in a regvar. Without such rules, the result would go in a scratch FREG, and we would need fmr to move it to the regvar. Also add a rule for pat sdl inreg($1)==reg_float with STACK, so we can unstack the value directly into the regvar, again without a scratch FREG and fmr. Edit util/ego/descr/powerpc.descr to tell ego about the new float regvars. This might not be working right; ego usually decides against using any float regvars, so ack -O1 (not running ego) uses the regvars, but ack -O4 (running ego) doesn't use the regvars. Beware that ack -mosxppc runs ego using powerpc.descr but -mlinuxppc and -mqemuppc run ego without a config file (since `8ef7c31`). I am testing powerpc.descr with a local edit to plat/linuxppc/descr to run ego with powerpc.descr there, but I did not commit my local edit.	2017-02-15 19:34:07 -05:00
George Koehler	13beb5e336	Document RELOLIS from commit `1bf58cf`. I hastily chose the name RELOLIS for this relocation type. If we want to rename it, we only need to edit these files: - h/out.h - mach/powerpc/as/mach5.c - util/amisc/ashow.c - util/led/ack.out.5 - util/led/relocate.c	2017-02-10 11:59:34 -05:00
George Koehler	1bf58cf51c	Add RELOLIS for PowerPC lis with ha16 or hi16. The new relocation type RELOLIS handles these instructions: lis RT, ha16[expr] == addis RT, r0, ha16[expr] lis RT, hi16[expr] == addis RT, r0, hi16[expr] RELOLIS stores a 32-bit value in the program text. In this value, the high bit is a ha16 flag, the next 5 bits are the target register RT, and the low bits are a signed 26-bit offset. The linker replaces this value with the lis instruction. The old RELOPPC relocated a ha16/lo16 or hi16/lo16 pair. The new RELOLIS relocates only a ha16 or hi16, so it is no longer necessary to have a matching lo16 in the next instruction. The disadvantage is that RELOLIS has only a signed 26-bit offset, not a 32-bit offset. Switch the assembler to use RELOLIS for ha16 or hi16 and RELO2 for lo16. The li32 instruction still uses the old RELOPPC relocation. This is not the same as my RELOPPC change from my recent mail to tack-devel (https://sourceforge.net/p/tack/mailman/message/35651528/). This commit is on a different branch. Here I am throwing away my RELOPPC change and instead trying RELOLIS.	2017-02-08 11:46:31 -05:00
George Koehler	a41b6f0458	Allow more PowerPC instructions in relocations. I need this for relocations in lis/lfd pairs. I add lfd along with addi, lfs, lha, stfs, stfd to the list.	2017-01-23 16:19:38 -05:00
George Koehler	f91bc2804d	Tune the installed manual pages. This commit slightly improves the formatting of the manuals. My OpenBSD machine uses mandoc(1) to format manuals. I check the manuals with `mandoc -T lint` and fix most of the warnings. I also make other changes where mandoc didn't warn me. roff(7) says, "Each sentence should terminate at the end of an input line," but we often forgot this rule. I insert some newlines after sentences that had ended mid-line. roff(7) also says that blank lines "are only permitted within literal contexts." I delete blank lines. This removes some extra blank lines from mandoc's output. If I do want a blank line in the output, I call ".sp 1" to make it in man(7). If I want a blank line in the source, but not the output, I put a plain dot "." so roff ignores it. Hyphens used for command-line options, like \-a, should be escaped by a backslash. I insert a few missing backslashes. mandoc warns if the date in .TH doesn't look like a date. Our manuals had a missing date or the RCS keyword "$Revision$". Git doesn't expand RCS keywords. I put in today's date, 2017-01-18. Some manuals used tab characters in filled mode. That doesn't work. I use .nf to turn off filled mode, or I use .IP in man(7) to make the indentation without a tab character. ack(1) defined a macro .SB but never used it, so I delete the definition. I also remove a call to the missing macro .RF. mandoc warns about empty paragraphs. I deleted them. mandoc also warned about these macro pairs in anm(1): .SM .B text The .SM did nothing because the .B text is on a different line. I changed each pair to .SB for small bold text. I make a few other small changes.	2017-01-18 23:02:30 -05:00
David Given	1a5c595a12	Merge pull request #45 from davidgiven/dtrg-fixups Add hi16[], ha16[], lo16[] support to the PowerPC assembler	2017-01-18 20:18:04 +01:00
David Given	d5a83fd73e	Clean up the led includes.	2017-01-18 19:55:56 +01:00
David Given	01a61e0708	Apply kernigh@'s fix to broken symbol tables in aelflod (via mailing list patch): ---snip--- The ELF spec at http://www.sco.com/developers/gabi/ says, "In each symbol table, all symbols with STB_LOCAL binding precede the weak and global symbols," and that sh_info is the index of the first non-local symbol. I was mixing local and global symbols and setting sh_info to zero. I also forgot to set the type of the .shstrtab section. ---snip---	2017-01-18 00:06:14 +01:00
David Given	b63a4513d5	Add missing header.	2017-01-15 12:04:47 +01:00
David Given	9a346c382d	Turns out Apple's hi16/ha16 exactly match my ha16/has16, so renamed accordingly. (Memo to self: read the docs before doing the work.)	2017-01-15 11:59:33 +01:00
David Given	f80acfe9f5	Signed vs unsigned lower halves of powerpc fixups are now handled by having two assembler directives, ha16() and has16(), for the upper half; has16() applies the sign adjustment. .powerpcfixup is now gone, as we generate the relocation in ha*() instead. Add special logic to the linker for undoing and redoing the sign adjustment when reading/writing fixups. Tests still pass.	2017-01-15 11:51:37 +01:00
David Given	14aab21204	Revert change; addis/ori requires different handling to addis/lwz due to ori's payload being unsigned while lwz's payload is signed.	2017-01-15 10:31:20 +01:00
David Given	8edbff9795	Add assembler support for fixing up arbitrary oris/addi pairs of instructions; this should allow oris/lwz constant value loads, which will save an opcode.	2017-01-15 00:15:01 +01:00
David Given	62022c6f6b	Don't print source file names during compilation (gcc stopped doing it years ago).	2017-01-08 00:16:35 +01:00
David Given	4b7fc5e233	Run through clang-format.	2017-01-08 00:15:23 +01:00
David Given	cca6171e55	Properly install man pages.	2017-01-07 22:38:30 +01:00
David Given	72766a02de	Fix typo in the descr file which was stopping -B from working. Add B documentation to the ack man page.	2017-01-04 13:28:40 +00:00
David Given	aeb9d4952d	Added an abmodules tool which detects B modules and generates an initialiser function for them (in C, unfortunately).	2017-01-03 18:54:13 +00:00
David Given	52f82a76b6	Don't crash when using the -u option to enter undefined symbols.	2016-12-28 23:49:55 +00:00
David Given	e50f4be710	Merge from default.	2016-12-26 19:44:48 +00:00
George Koehler	7e39a821f4	Update aelflod.1 to describe options, mention symbol table.	2016-12-13 17:22:45 -05:00
George Koehler	5cf2a68438	Teach aelflod to write the ELF symbol table. Convert each ack.out symbol to ELF, preserving its name and value, and some but not all other symbol info. The ELF symbol table comes with ELF section headers. If the input file has no symbols (ack -Rled-s), then the output file has no ELF section headers. Append the symbol table and section headers to the ELF file. Linux ignores this appended info when it execs the file. The info might pollute the last page of the ELF segment, but Linux clears this pollution. Tools like nm and gdb can read the ELF symbol table. I include code to translate debugger symbols to an ELF .stab section. I did not test this code, because I did not find a platform that can put S_STB symbols in the ack.out file.	2016-12-13 16:41:22 -05:00
George Koehler	8280ca8745	In aslod, remove some unused m68k2 stuff.	2016-12-13 16:02:38 -05:00
George Koehler	2874806048	Tweak man syntax in aelflod.1 and aslod.1 Put end of sentence at end of line. This is roff(7) syntax so the formatter can make spacing between sentences. Use the macro .PP to break paragraphs. Use bold for the command name in SYNOPSIS, to match other ack manuals.	2016-12-13 15:54:38 -05:00
David Given	3cce93586f	Merge branch 'pr-osx' of https://github.com/kernigh/ack into kernigh-pr-osx	2016-12-05 20:14:54 +01:00
David Given	07e654d6e3	Merge from default.	2016-12-02 00:25:50 +01:00
David Given	8c99e2b7ad	Run all tests, even the ones which fail, and emit a test summary right at the end of the build (and fail then).	2016-12-01 23:03:30 +01:00
David Given	b7de58e34e	Fix signal unsafety in testdriver.c.	2016-12-01 22:05:22 +01:00
David Given	90547157b4	Merge from default.	2016-11-30 22:07:29 +01:00
David Given	8ce3382996	Merge from default.	2016-11-30 22:04:11 +01:00
David Given	d94c326465	Merge pull request #18 from kernigh/pr-util-ack util/ack: use DEFAULT_PLATFORM, allow non-ASCII filenames, other changes	2016-11-30 22:00:55 +01:00
George Koehler	8ef7c31089	Write a powerpc.descr for ego and use it with osxppc. No change to linuxppc and qemuppc. They continue to run ego without any descr file. I copied m68020.descr to powerpc.descr and changed some numbers. My numbers are guesses; I know little about PowerPC cycle counts, and almost nothing about ego. This powerpc.descr causes most of the example programs to shrink in size (without descr -> with descr): 65429 -> 57237 hilo_b.osxppc -8192 36516 -> 32420 hilo_c.osxppc -4096 55782 -> 51686 hilo_mod.osxppc -4096 20096 -> 20096 hilo_p.osxppc 0 8813 -> 8813 mandelbrot_c.osxppc 0 93355 -> 89259 paranoia_c.osxppc -4096 92751 -> 84559 startrek_c.osxppc -8192 (Each file has 2 Mach segments, then a symbol table. Each segment takes a multiple of 4096 bytes. When the code shrinks, we lose a multiple of 4096 bytes.) I used "ack -mosxppc -O6 -c.so" to examine the assembly code for hilo.mod and mandelbrot.c, both without and with descr. This reveals optimizations made only with descr, from 2 ego phases: SP (stack pollution) and RA (register allocation). In hilo.mod, SP deletes some instructions that remove items from the stack. These items get removed when the function returns. In both hilo.mod and mandelbrot.c, RA moves some values into local variables, so ncg can make them into register variables. This shrinks code size, probably because register variables get preserved across function calls. More values stay in registers, and ncg emits shorter code. I believe that the ego descr file uses (time,space) tuples but the ncg table uses (space,time) tuples. This is confusing. Perhaps I am wrong, and some or all tuples are backwards. My time values are the cycle counts in latency from the MPC7450 Reference Manual (but not including complications like "store serialization"). In powerpc.descr, I give the cost for saving and restoring registers as if I was using chains of stw and lwz instructions. Actually ncg uses single stmw and lmw instructions with at least 2 instructions. The (time,space) for stmw and lmw would be much less than the (time,space) for chains of stw and lwz. But this ignores the pipeline of the MPC7450. The chains of stw and lwz may run faster than stmw and lmw in the pipeline, because the throughput may be better than the latency. By using the wrong values for (time,space), I'm trying to tell ego that stmw and lmw are not better than chains of stw and lwz.	2016-11-30 15:29:19 -05:00
David Given	960584c0f3	Replace the hacky and broken pipeline in testdriver.sh with a custom-written tool in C; much more robust and easier to understand, as well as avoiding the dependency on timeout (which isn't Posix).	2016-11-29 20:59:43 +01:00
George Koehler	ecdfb61c9d	Merge branch 'default' into kernigh-osx This brings in David Given's PowerPC changes, including the addition of the modern code generator (mcg) for PowerPC. Resolve minor conflicts in top build.lua and util/led/main.c	2016-11-28 16:20:56 -05:00
David Given	5bce5fc4da	Change the extension used by Basic files for .b to .bas, to avoid conflicts with B.	2016-11-27 20:38:33 +01:00
George Koehler	88c2ea63aa	Use uint32_t in util/led/main.c This uses uint32_t for the base, file offset, and alignment of each section, to be consistent with the usage of uint32_t in h/out.h Also declare setbit() as static.	2016-11-20 11:38:16 -05:00
David Given	d5328492d7	Better handling of float conversions; more tests; converting to unsigned ints works now.	2016-11-20 11:27:40 +01:00
David Given	d31bc6a3f9	Made csa and csb work with mcg; adjust the libem functions and the corresponding invocation in the ncg table so the same helpers can be used for both mcg and ncg. Add a new IR opcode, FARJUMP, which jumps to a helper function but saves volatile registers.	2016-11-19 10:55:41 +01:00
George Koehler	486c516242	Stop trying to remove core dumps. unlink("core") doesn't work with OpenBSD, where core dumps have names like "ncg.core". Users who don't want core dumps can turn them off with "ulimit -c 0" in sh(1). Then the system doesn't write a core dump. That's better than writing core then unlinking it.	2016-11-15 11:58:13 -05:00
George Koehler	07b04c64c8	Let ack(1) compile files with non-ASCII names. This commit changes how ack(1) parses backslashes in its descr files. Before this commit, ack set the high bit of each character escaped by a backslash, and later cleared all high bits in command arguments, but this lost the high bits in non-ASCII filenames. After this commit, ack keeps backslashes in strings while processing them. Functions scanvars(), scanexpr(), doassign(), unravel(), addargs() now understand backslashes. I remove from ack_basename() the warning about non-ASCII characters. This commit makes some incompatible changes for backslashes in descr files. None of our descr files uses backslashes, except for those backslashes that continue lines, and there are no changes for those backslashes. The problem with non-ASCII filenames had its cause in a feature that we weren't using. With this commit, ack now understands backslashes after the = sign in both "var NAME=value" and "mapflag -flag NAME=value". Before, ack never scanned backslashes in "var" lines, so "var A=\{B}" failed to prevent expansion of B. Now it does. Before, ack did scan for backslashes in the "-flag NAME=" part of "mapflag" lines. Now it doesn't, so it is no longer possible to map a flag that contains a literal space, tab, or star "*". I removed the expansion of "{{" to "{". One can use "\{" for a literal "{", and "\{" now works in "var" lines. Before and now, ack never expanded "{" in flags for "mapflag", so the correct way to map a literal flag "-{" remains "mapflag -{ ...", not "mapflag -{{ ...". (The other way "mapflag -\{ ..." stops working with this commit.) Backslashes in strange places, like "{NA\ME}", probably have different behavior now. Backslashes in "program" lines now work. Before, ack scanned for backslashes there but forgot to clear the high bits later. Escaping < or > as \< or \> now works, and prevents substitution of the input or output file paths. Before, ack only expanded the first < or > in each argument. Now, it expands every unescaped < or > in an argument, but this is an accident of how I rewrote the code. I don't suggest to put more than one each of < or > in a command. The code no longer optimizes away its recursive calls when the argument is "<". The code continues to set or clear the high bit NO_SCAN on the first characters of flags. This doesn't seem to be a problem, because flags usually begin with an ASCII hyphen '-'.	2016-11-15 11:33:35 -05:00
David Given	408b69b17d	aelflod and aslod now default to not showing the memory dump.	2016-11-13 20:50:23 +01:00
George Koehler	fafc8a0b8a	Don't retry fork() in a loop. If fork() fails, then report a fatal error. Don't spin the cpu retrying fork() until it succeeds. It can fail when we reach a limit on the number of processes. Spinning on the cpu would slow down other processes when we want them to exit. This would get bad if we had a parallel build with multiple ack processes spinning.	2016-11-13 12:45:01 -05:00
George Koehler	e617f42503	Don't print a string after possibly freeing it. new= newvar(name) takes ownership of the string and might free its memory. Don't print name. Do print new->v_name. Also #include <string.h> for strcmp().	2016-11-13 12:02:35 -05:00
George Koehler	65278bc9a9	In util/ack, add prototypes and static to functions. Declare most functions before using them. I declare some functions in ack.h and some in trans.h (because trans.h declares type trf). I leave declarations of scanb() and scanvars() in .c files because they need type growstring. (I can't #include "grows.h" in another header file as long as grows.h doesn't guard against multiple inclusion.) Functions used within one file become static. I remove a few tiny functions. I move a few functions or declarations to more convenient places. Some functions now return void instead of a garbage int. I feel that keyword "register" is obsolete, so I removed it from where I was editing. This commit doesn't touch mktables.c	2016-11-12 21:12:45 -05:00
George Koehler	7b3f870f63	Don't build mktables.c in the ack binary. It only worked by accident because main() in main.c was found before main() in mktables.c. Also add build dependencies on the local *.h files.	2016-11-11 17:06:25 -05:00
David Given	84ee75ec07	Merge from default.	2016-11-11 20:17:54 +01:00
David Given	d82df74a7a	Rename addr_t to address_t to avoid clashes with the system addr_t.	2016-11-11 20:17:10 +01:00
David Given	fd91851005	Add enough return types to the K&R C that the ACK builds (on Linux) using clang now.	2016-11-10 22:04:18 +01:00
David Given	a8c4dac67c	Merge from default (merging in George Koehler's PowerPC changes).	2016-10-29 22:40:40 +02:00

1 2 3 4 5 ...

1688 commits