6502bench

mirror of https://github.com/fadden/6502bench.git synced 2024-12-01 22:50:35 +00:00

Author	SHA1	Message	Date
Andy McFadden	ee58d9e803	Data Bank Register management, part 3 Added a "fake" assembler pseudo-op for DBR changes. Display entries in line list. Added entry to double-click handler so that you can double-click on a PLB instruction operand to open the data bank editor.	2020-07-09 16:52:23 -07:00
Andy McFadden	3820bfee8b	Set initial focus to appropriate field for project properties If you double-click a project symbol declaration, the symbol editor opens. I found that I was double-clicking on the comment field and typing with the expectation that the comment would be updated, but it was actually setting the initial focus to the label field. With this change the symbol editor will focus the label, value, or comment field based on which column was double-clicked. The behavior for Actions > Edit Project Symbol and other paths to the symbol editor are unchanged. Also, disabled a wayward assert.	2020-04-23 11:01:07 -07:00
Andy McFadden	071adb8e95	Two changes to "dense hex" bulk data formatting (1) Added an option to limit the number of bytes per line. This is handy for things like bitmaps, where you might want to put (say) 3 or 8 bytes per line to reflect the structure. (2) Added an application setting that determines whether the screen listing shows Merlin/ACME dense hex (20edfd) or 64tass/cc65 hex bytes ($20,$ed,$fd). Made the setting part of the assembler-driven display definitions. Updated 64tass+cc65 to use ".byte" as their dense hex pseudo-op, and to use the updated formatter code. No changes to regression test output. (Changes were requested in issue #42.) Also, added a resize gripper to the bottom-right corner of the main window. (These seem to have generally fallen out of favor, but I like having it there.)	2019-12-10 17:41:00 -08:00
Andy McFadden	88c56616f7	Label rework, part 5 Implemented assembly source generation of non-unique local labels. The new 2023-non-unique-labels test exercises various edge cases (though we're still missing local variable interaction). The format of uniquified labels changed slightly, so the expected output of 2012-label-localizer needed to be updated. This changes the "no opcode mnemonics" and "mask leading underscores" functions into integrated parts of the label localization process.	2019-11-17 16:05:51 -08:00
Andy McFadden	4e08810278	Finish removal of "disable label localizer" feature The label localizer is now always on. The regression tests turned it off by default, but that's no longer allowed, so the generated output has changed for many of them. The tests themselves were not altered.	2019-11-16 17:15:03 -08:00
Andy McFadden	be65f280a3	Minor tweaks - Renamed "strip label prefix/suffix" to "omit label prefix/suffix". - Changed a Merlin operand workaround so it doesn't apply to code that is explicitly not in bank zero. - Changed {addr}/{const} annotations on project/platform symbol equates so they line up a little better on screen and in exported sources.	2019-11-15 16:24:07 -08:00
Andy McFadden	5dd7576529	Label rework, part 2 Continue development of non-unique labels. The actual labels are still unique, because we append a uniquifier tag, which gets added and removed behind the scenes. We're currently using the six-digit hex file offset because this is only used for internal address symbols. The label editor and most of the formatters have been updated. We can't yet assemble code that includes non-unique labels, but older stuff hasn't been broken. This removes the "disable label localization" property, since that's fundamentally incompatible with what we're doing, and adds a non- unique label prefix setting so you can put '@' or ':' in front of your should-be-local labels. Also, fixed a field name typo.	2019-11-12 17:44:51 -08:00
Andy McFadden	4d079c8d14	Label rework, part 1 This adds the concept of label annotations. The primary driver of the feature is the desire to note that sometimes you know what a thing is, but sometimes you're just taking an educated guess. Instead of writing "high_score_maybe", you can now write "high_score?", which is more compact and consistent. The annotations are stripped off when generating source code, making them similar to Notes. I also created a "Generated" annotation for the labels that are synthesized by the address table formatter, but don't modify the label for them, because there's not much need to remind the user that "T1234" was generated by algorithm. This also lays some of the groundwork for non-unique labels.	2019-11-08 21:02:15 -08:00
Andy McFadden	b6e571afc2	Correctly handle embedded instruction edge case This began with a change to support "BRK <operand>" in cc65. The assembler only supports this for 65816 projects, so we detect that and enable it when available. While fiddling with some test code an assertion fired. This revealed a minor issue in the code analyzer: when overwriting inline data with instructions, we weren't resetting the format descriptor. The code that exercises it, which requires two-byte BRKs and an inline BRK handler in an extension script, has been added to test 2022-extension-scripts. The new regression test revealed a flaw in the 64tass code generator's character encoding scanner that caused it to hang. Fixed.	2019-10-19 17:28:45 -07:00
Andy McFadden	cd23580cc5	Add junk/align directives Sometimes there's a bunch of junk in the binary that isn't used for anything. Often it's there to make things line up at the start of a page boundary. This adds a ".junk" directive that tells the disassembler that it can safely disregard the contents of a region. If the region ends on a power-of-two boundary, an alignment value can be specified. The assembly source generators will output an alignment directive when possible, a .fill directive when appropriate, and a .dense directive when all else fails. Because we're required to regenerate the original data file, it's not always possible to avoid generating a hex dump.	2019-10-18 21:00:28 -07:00
Andy McFadden	dfd5bcab1b	Optionally treat BRKs as two-byte instructions Early data sheets listed BRK as one byte, but RTI after a BRK skips the following byte, effectively making BRK a 2-byte instruction. Sometimes, such as when diassembling Apple /// SOS code, it's handy to treat it that way explicitly. This change makes two-byte BRKs optional, controlled by a checkbox in the project settings. In the system definitions it defaults to true for Apple ///, false for all others. ACME doesn't allow BRK to have an arg, and cc65 only allows it for 65816 code (?), so it's emitted as a hex blob for those assemblers. Anyone wishing to target those assemblers should stick to 1-byte mode. Extension scripts have to switch between formatting one byte of inline data and formatting an instruction with a one-byte operand. A helper function has been added to the plugin Util class. To get some regression test coverage, 2022-extension-scripts has been configured to use two-byte BRK. Also, added/corrected some SOS constants. See also issue #44.	2019-10-09 14:55:56 -07:00
Andy McFadden	824add17e8	Remap labels that use opcode mnemonics In a recent survey, three out of four cross assemblers surveyed recommended not using opcode mnemonics to their patients who use labels. We now remap labels like "AND" and "jmp", using the label map that's part of the label localizer. We skip the step for Merlin 32, which is perfectly happy to assemble "JMP JMP JMP". Also, fixed a bug in MaskLeadingUnderscores that could hang the source generator thread.	2019-09-20 15:29:34 -07:00
Andy McFadden	b74630dd5b	Work around two assembler issues Most assemblers end local label scope when a global label is encountered. cc65 takes this one step further by ending local label scope when constants or variables are defined. So, if we have a variable table with a nonzero number of entries, we want to create a fake global label at that point to end the scope. Merlin 32 won't let you write " LDA #',' ". For some reason the comma causes an error. IGenerator now has a "tweak operand format" interface that lets us fix that.	2019-09-20 14:05:17 -07:00
Andy McFadden	3353819a62	Change the way the "add padded string" functions work The functions started by trying to pad a column out to a width, then changed to pad things to a certain length. What they really should be doing is padding the start of an entry to a specified column. This is much more natural and avoids a trim operation. The only change to the output is to ORG statements from the HTML exporter, which are now formatted correctly.	2019-09-17 22:02:05 -07:00
Andy McFadden	14b215b76d	Implement local variables for cc65 I'd apparently overlooked the ".set" directive, which seems to do exactly what we need.	2019-09-01 18:14:39 -07:00
Andy McFadden	d542a809f8	Implement local variables for ACME Unlike 64tass and Merlin, which allow you to redefine symbols, ACME uses "zones" that provide scope for local variables. This means that, at the point of a local variable table definition, we have to start a new zone and output the full set of active symbols, not just the newly-defined ones. (If you set the "clear previous" flag in the LvTable there's no difference.) We could do a bit better by only outputting the symbols that are actually used within the zone, similar to what we do for global project/platform symbols, but that's a bunch of work for questionable benefit.	2019-09-01 10:55:19 -07:00
Andy McFadden	e82339573f	Add VarDirective to PseudoOpNames Also, rearranged the pseudo-op app settings XAML to be a bit easier to maintain.	2019-08-29 12:14:47 -07:00
Andy McFadden	32d1147eec	Improve multi-encoding output in 64tass Previously, we used the default character encoding from the project properties to determine how strings and character constants in the entire source file should be encoded. Now we switch between encodings as needed. The default character encoding is no longer relevant. High ASCII is now an actual encoding, rather than acting like ASCII that sometimes doesn't work. Because we can do high ASCII character operands with "\| $80", we don't output a .enc to switch from ASCII to high ASCII unless we need to generate a string. (If we're already in high ASCII mode, the "\| $80" isn't required but won't hurt anything.) We now do a scan up front to see if ASCII or high ASCII is needed, and only output the .cdefs for the encodings that are actually used. The only gap in the matrix is high ASCII DCI strings -- the ".shift" pseudo-op rejects text if the string doesn't start with the high bit clear.	2019-08-21 13:46:05 -07:00
Andy McFadden	4902b89cf8	Various improvements The PseudoOpNames class is increasingly being used in situations where mutability is undesirable. This change makes instances immutable, eliminating the Copy() method and adding a constructor that takes a Dictionary. The serialization code now operates on a Dictionary instead of the class properties, but the JSON encoding is identical, so this doesn't invalidate app settings file data. Added an equality test to PseudoOpNames. In LineListGen, don't reset the line list if the names haven't actually changed. Use a table lookup for C64 character conversions. I figure that should be faster than multiple conditionals on a modern x64 system. Fixed a 64tass generator issue where we tried to query project properties in a call that might not have a project available (specifically, getting FormatConfig values out of the generator for use in the "quick set" buttons for Display Format). Fixed a regression test harness issue where, if the assembler reported success but didn't actually generate output, an exception would be thrown that halted the tests. Increased the width of text entry fields on the Pseudo-Op tab of app settings. The previous 8-character limit wasn't wide enough to hold ACME's "!pseudopc". Also, use TrimEnd() to remove trailing spaces (leading spaces are still allowed). In the last couple of months, Win10 started stalling for a fraction of a second when executing assemblers. It doesn't do this every time; mostly it happens if it has been a while since the assembler was run. My guess is this has to do with changes to the built-in malware scanner. Whatever the case, we now change the mouse pointer to a wait cursor while updating the assembler version cache.	2019-08-17 11:30:42 -07:00
Andy McFadden	84d3146903	Update source generators to recognize C64 strings For the most part this means explicitly dumping them as hex, though ACME gets to exercise its !pet and !scr operators.	2019-08-15 21:33:10 -07:00
Andy McFadden	beb1024550	Define and use "delimiter sets" A delimiter definition is four strings (prefix, open, close, suffix) that are concatenated with the character or string data to form an operand. A delimiter set is a collection of delimiter definitions, with separate entries for each character encoding. This is a convenient way to configure Formatter objects, import and export data from the app settings file, and manage the UI needed to allow the user to customize how things look. The full set of options didn't fit on the first app settings tab, so there's now a separate tab just for specifying character and string delimiters. (This might be overkill, but there are various plausible scenarios that make use of it.) The delimiters for on-screen display of strings can now be configured.	2019-08-14 16:10:04 -07:00
Andy McFadden	5889f45737	Replace on-screen string operand formatting The previous functions just grabbed 62 characters and slapped quotes on the ends, but that doesn't work if we want to show strings with embedded control characters. This change replaces the simple formatter with the one used to generate assembly source code. This increases the cost of refreshing the display list, so a cache will need to be added in a future change. Converters for C64 PETSCII and C64 Screen Code have been defined. The results of changing the auto-scan encoding can now be viewed. The string operand formatter was using a single delimiter, but for the on-screen version we want open-quote and close-quote, and might want to identify some encodings with a prefix. The formatter now takes a class that defines the various parts. (It might be worth replacing the delimiter patterns recently added for single-character operands with this, so we don't have two mechanisms for very nearly the same thing.) While working on this change I remembered why there were two kinds of "reverse" in the old Merlin 32 string operand generator: what you want for assembly code is different from what you want on screen. The ReverseMode enum has been resurrected.	2019-08-13 17:52:58 -07:00
Andy McFadden	f33cd7d8a6	Replace character operand output method The previous code output a character in single-quotes if it was standard ASCII, double-quotes if high ASCII, or hex if it was neither of those. If a flag was set, high ASCII would also be output as hex. The new system takes the character value and an encoding identifier. The identifier selects the character converter and delimiter pattern, and puts the two together to generate the operand. While doing this I realized that I could trivially support high ASCII character arguments in all assemblers by setting the delimiter pattern to "'#' \| $80". In FormatDescriptor, I had previously renamed the "Ascii" sub-type "LowAscii" so it wouldn't be confused, but I dislike filling the project file with "LowAscii" when "Ascii" is more accurate and less confusing. So I switched it back, and we now check the project file version number when deciding what to do with an ASCII item. The CharEncoding tests/converters were also renamed. Moved the default delimiter patterns to the string table. Widened the delimiter pattern input fields slightly. Added a read- only TextBox with assorted non-typewriter quotes and things so people have something to copy text from.	2019-08-11 22:11:00 -07:00
Andy McFadden	975b62db6b	Treat low and high ASCII as two distinct formats We've been treating ASCII strings and instruction/data operands as ambiguous, resolving low vs. high when generating output for the display or assembler. This change splits it into two separate formats, simplifying output generation. The UI will continue to treat low/high ASCII as as single thing, selecting the format appropriately based on the data. There's no reason to have two radio buttons that are never both enabled. The data operand string functions need some additional work, but that overlaps substantially with the upcoming PETSCII changes, so for now all strings set by the data operand editor are low ASCII. The file format has changed again, but since there hasn't been a release since the previous change, I'm leaving the file format at v2. Code has been added to resolve the ASCII mode when loading a v1 project file. This removes some complexity from the assembly code generators.	2019-08-10 14:59:24 -07:00
Andy McFadden	dae76d9b45	Rework string operand formatting This generalizes the string pseudo-operand formatter, moving it into the Asm65 library. The assembly source generators have been updated to use it. This makes the individual generators simpler, and by virtue of avoiding "test runs" should make them slightly faster. This also introduces byte-to-character converters, though we're currently still only supporting low/high ASCII. Regression test output is unchanged.	2019-08-09 17:46:33 -07:00
Andy McFadden	835c1c7fe2	Reverse position on '#' in block move operands During a discussion with the cc65 developers, I became convinced that generating "MVN $01,$02" is wrong, and "MVN #$01,#$02" is correct. 64tass, cc65, and Merlin 32 all accept this syntax; only ACME does not. Operands without a leading '#' should be treated as 24-bit values, and have the bank byte extracted. This change updates the on-screen display and assembled output to include the '#'. The ACME generator uses a Quirk to suppress the hash mark. (It doesn't currently accept values larger than 8 bits, so there's no ambiguity.)	2019-08-08 13:02:01 -07:00
Andy McFadden	0d0854bda7	Change the way string formats are defined We used to use type="String", with the sub-type indicating whether the string was null-terminated, prefixed with a length, or whatever. This didn't leave much room for specifying a character encoding, which is orthogonal to the sub-type. What we actually want is to have the type specify the string type, and then have the sub-type determine the character encoding. These sub-types can also be used with the Numeric type to specify the encoding of character operands. This change updates the enum definitions and the various bits of code that use them, but does not add any code for working with non-ASCII character encodings. The project file version number was incremented to 2, since the new FormatDescriptor serialization is mildly incompatible with the old. (Won't explode, but it'll post a complaint and ignore the stuff it doesn't recognize.) While I was at it, I finished removing DciReverse. It's still part of the 2005-string-types regression test, which currently fails because the generated source doesn't match.	2019-08-07 16:19:13 -07:00
Andy McFadden	71badf2359	Update for cc65 v2.18 WDM <arg> now works. MVN/MVP are still broken. Correct code is generated for whichever version of the assembler is configured. Regression tests updated for new version. Also, fixed a UI bug where manual edits to the assembler path were being ignored.	2019-08-04 13:38:25 -07:00
Andy McFadden	1ad9caa783	First pass at ACME support I managed to work around most of the quirks, but there's still an issue with 65816 code. Also, enabled word wrapping in the AsmGen text boxes.	2019-08-03 20:54:07 -07:00
Andy McFadden	98914e9f80	Treat BRK as a 1-byte instruction The 65816 definition makes it a two-byte instruction, like COP. On the 6502 it acted like a two-byte instruction, but in practice very few assemblers treat it that way. Very few humans, for that matter. So it's now treated as a single byte instruction, with the following byte encoded as a data value.	2019-08-02 17:21:50 -07:00
Andy McFadden	c64f72d147	Move WPF code from SourceGenWPF to SourceGen	2019-07-20 13:28:37 -07:00
Andy McFadden	e3906e021b	Move WinForms code to SourceGenWF	2019-07-20 13:02:54 -07:00
Andy McFadden	2065f4ef9e	Attempt to generate segment names for cc65 This worked, sort of. The problem is that SourceGen will revert to hex output in certain situations, such as a broken symbolic reference. There happens to be one in the ZIPPY example, and it's on a relative branch. The goal with the segment stuff is to allow cc65 to treat the source as relocatable code. In that context, a relative branch to an absolute address doesn't make any sense, so the assembler reports a range error. We don't currently have a mechanism that guarantees no references are broken (and no affordance for finding them), so we can't make this mode the default yet. Instead, we continue to use the generic config, but generate the correct set of lines as comments. (issue #39)	2018-11-18 15:11:29 -08:00
Andy McFadden	17f0faa845	Add linker config scripts to cc65 generator output The system configuration you get with "-t none" works for smaller files but fails for larger ones. This updates the generator to produce a source file and linker script pair. (I kinda saw this one coming -- it's why the gen/asm dialog has a combo box for the file preview -- so it didn't require that much work.) This currently generates a fixed script for a generic system with 64KiB of RAM, using .ORGs to set the addresses as before. With this change, assembling a file with 65536 NOPs succeeds. (issue #39)	2018-11-18 14:28:44 -08:00
Andy McFadden	5b1dde290a	Show assembler options in header comment	2018-11-03 14:02:52 -07:00
Andy McFadden	a88c746419	Work around cc65 single-pass behavior The cc65 assembler runs in a single pass, which means forward address references default to 16 bits. For zero-page references we have to add an explicit width disambiguator. (This is an unusual situation that only occurs if you have a zero-page .ORG in the file after code that references it.) With this change, 2014-label-dp passes, and no other regression tests were affected. (issue #40)	2018-11-02 15:32:54 -07:00
Andy McFadden	c80be07f73	Work around Merlin 32 instruction parsing bug The 2014-label-dp test now passes. Prior regression tests are unaffected. Also, renamed an IGenerator interface to more accurately reflect its role. (issue #37)	2018-11-02 13:49:27 -07:00
Andy McFadden	7aa3e4dbcd	Show "assembling" when assembling Merlin 32 is slow enough with a 64K data file that you have enough time to read the text.	2018-10-30 16:41:56 -07:00
Andy McFadden	a8af7e8794	Improve the "common" expression formatter To avoid confusing the assembler, expressions with a leading parenthesis like "(foo & $ffff) + 1" are prefixed with a "0+". This is not necessary if the operand begins with a '#'. (issue #16)	2018-10-26 15:45:39 -07:00
Andy McFadden	61914c8f79	Progress toward 64tass expression support Gave cc65 its own expression generator, as the precedence table seems atypical if not unique. Configured 64tass to use the "simple" expression mode. Added some operations on a 32-bit constant to 2007-labels-and-symbols to exercise the current worst-case expression (shift + AND + add). Tweaked the Merlin expression generator to handle it. (issue #16)	2018-10-24 13:17:03 -07:00
Andy McFadden	f7e5cf2f45	Progress toward 64tass support Most tests pass, but 2007-labels-and-symbols fails because the expressions recognized by 64tass don't match up with either of the other assemblers. This is currently using a workaround for the local label syntax. 64tass uses '_' as the prefix, which is unfortunate since SourceGen explicitly allowed underscores in labels. (So does 64tass for that matter, but it treats labels specially when the '_' comes first.) We will need to rename any non-local user labels that start with '_'. (issue #16)	2018-10-23 20:08:01 -07:00
Andy McFadden	ab9287fef8	Progress toward new assembler configuration Changed the "quick config" buttons for the asm config and pseudo-op tabs into a drop-list and "set" button. The default values for each assembler are now defined in the Asm*.cs file, rather than in the settings code.	2018-10-21 16:36:48 -07:00
Andy McFadden	9aabd988a8	Progress toward new assembler configuration Use configured column widths when generating output. The regression test always uses the assembler-preferred default widths.	2018-10-20 21:24:28 -07:00
Andy McFadden	4f9af30455	Progress toward new assembler configuration Rather than have each assembler get its own app config string for the cross-assembler executable, we now have a collection of per- assembler config items, of which the executable path name is one member. The Asm Config tab has an auto-generated pop-up to select the assembler. The per-assembler settings block is serialized with the rather unpleasant JSON-in-JSON approach, but nobody should have to look at it. This also adds assembler-specific column widths to the settings dialog, though they aren't actually used yet.	2018-10-20 20:35:32 -07:00
Andy McFadden	f81c534d25	Merge Gen* and Asm* source files Each supported assembler has an IGenerator interface and an IAssembler interface. They're still two separate classes, but now both are implemented in the same source file. (They'll probably stay separate classes, since the two have little interaction.) I'm keeping the "Asm*" filename. Seems the more natural fit. Also, changed AssemblerInfo to try to get all assembler-specific stuff into a single table.	2018-10-17 13:50:28 -07:00
Andy McFadden	2c6212404d	Initial file commit	2018-09-28 10:05:11 -07:00

46 Commits