ORCA-C

mirror of https://github.com/byteworksinc/ORCA-C.git synced 2024-05-28 12:41:35 +00:00

Author	SHA1	Message	Date
Stephen Heumann	99a10590b1	Avoid out-of-range branches around asm code using dcl directives. The branch range calculation treated dcl directives as taking 2 bytes rather than 4, which could result in out-of-range branches. These could result in linker errors (for forward branches) or silently generating wrong code (for backward branches). This patch now treats dcb, dcw, and dcl as separate directives in the native-code layer, so the appropriate length can be calculated for each. Here is an example of code affected by this: int main(int argc, char *argv) { top: if (!argc) { / this caused a linker error / asm { dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 dcl 0 } goto top; / this generated bad code with no error */ } }	2022-10-13 18:00:16 -05:00
Stephen Heumann	19683706cc	Do not optimize code from asm statements. Previously, the assembly-level optimizations applied to code in asm statements. In many cases, this was fine (and could even do useful optimizations), but occasionally the optimizations could be invalid. This was especially the case if the assembly involved tricky things like self-modifying code. To avoid these problems, this patch makes the assembly optimizers ignore code from asm statements, so it is always emitted as-is, without any changes. This fixes #34.	2022-10-12 22:03:37 -05:00
Stephen Heumann	995ded07a5	Always treat "struct T;" as declaring the tag within the current scope. A declaration of this exact form always declares the tag T within the current scope, and as such makes this "struct T" a distinct type from any other "struct T" type in an outer scope. (Similarly for unions.) See C17 section 6.7.2.3 p7 (and corresponding places in all other C standards). Here is an example of a program affected by this: struct S {char a;}; int main(void) { struct S; struct S sp; struct S {long b;} s; sp = &s; sp->b = sizeof(sp); return s.b; }	2022-10-04 18:45:11 -05:00
Stephen Heumann	05ecf5eef3	Add option to use the declared type for float/double/comp params. This differs from the usual ORCA/C behavior of treating all floating-point parameters as extended. With the option enabled, they will still be passed in the extended format, but will be converted to their declared type at the start of the function. This is needed for strict standards conformance, because you should be able to take the address of a parameter and get a usable pointer to its declared type. The difference in types can also affect the behavior of _Generic expressions. The implementation of this is based on ORCA/Pascal, which already did the same thing (unconditionally) with real/double/comp parameters.	2022-09-18 21:16:46 -05:00
Stephen Heumann	4e76f62b0e	Allow additional letters in identifiers. The added characters are accented roman letters that were added to the Mac OS Roman character set at some time after it was first defined. Some IIGS fonts include them, although others do not.	2022-08-01 19:59:49 -05:00
Stephen Heumann	1177ddc172	Tweak release notes. The "known issue" about not issuing required diagnostics is removed because ORCA/C has gotten significantly better about that, particularly if strict type checking is enabled. There are still probably some diagnostics that are missed, but it is no longer a big enough issue to be called out more prominently than other bugs.	2022-07-19 20:38:13 -05:00
Stephen Heumann	6e3fca8b82	Implement strict type checking for enum types. If strict type checking is enabled, this will prohibit redefinition of enums, like: enum E {a,b,c}; enum E {x,y,z}; It also prohibits use of an "enum E" type specifier if the enum has not been previously declared (with its constants). These things were historically supported by ORCA/C, but they are prohibited by constraints in section 6.7.2.3 of C99 and later. (The C90 wording was different and less clear, but I think they were not intended to be valid there either.)	2022-07-19 20:35:44 -05:00
Stephen Heumann	d576f19ede	Remove trailing whitespace in release notes. (No substantive changes.)	2022-07-18 21:45:55 -05:00
Stephen Heumann	6d07043783	Do not treat uses of enum types from outer scopes as redeclarations. This affects code like the following: enum E {a,b,c}; int main(void) { enum E e; struct E {int x;}; /* or: enum E {x,y,z}; */ } The line "enum E e;" should refer to the enum type declared in the outer scope, but not redeclare it in the inner scope. Therefore, a subsequent struct, union, or enum declaration using the same tag in the same scope is acceptable.	2022-07-18 21:34:29 -05:00
Stephen Heumann	2cbcdc736c	Allow the same identifier to be used as a typedef and an enum tag. This should be allowed (because they are in separate name spaces), but was not. This affected code like the following: typedef int T; enum T {a,b,c};	2022-07-18 18:33:54 -05:00
Stephen Heumann	6bfd491f2a	Update release notes.	2022-07-14 18:40:59 -05:00
Stephen Heumann	7b0dda5a5e	Fix a flawed optimization. The optimization could turn an unsigned comparison "x <= 0xFFFF" into "x < 0". Here is an example affected by this: int main(void) { unsigned i = 1; return (i <= 0xffff); }	2022-07-10 22:25:55 -05:00
Stephen Heumann	7898c619c8	Fix several cases where a condition might not be evaluated correctly. These could occur because the code for certain operations was assumed to set the z flag based on the result value, but did not actually do so. The affected operations were shifts, loads or stores of bit-fields, and ? : expressions. Here is an example showing the problem with a shift: #pragma optimize 1 int main(void) { int i = 1, j = 0; return (i >> j) ? 1 : 0; } Here is an example showing the problem with a bit-field load: struct { signed int i : 16; } s = {1}; int main(void) { return (s.i) ? 1 : 0; } Here is an example showing the problem with a bit-field store: #pragma optimize 1 struct { signed int i : 16; } s; int main(void) { return (s.i = 1) ? 1 : 0; } Here is an example showing the problem with a ? : expression: #pragma optimize 1 int main(void) { int a = 5; return (a ? (a<<a) : 0) ? 0 : 1; }	2022-07-07 18:26:37 -05:00
Stephen Heumann	497e5c036b	Use new 16-bit unsigned multiply routine that complies with C standards. This changes unsigned 16-bit multiplies to use the new ~CUMul2 routine in ORCALib, rather than ~UMul2 in SysLib. They differ in that ~CUMul2 gives the low-order 16 bits of the true result in case of overflow. The C standards require this behavior for arithmetic on unsigned types.	2022-07-06 22:22:02 -05:00
Stephen Heumann	f6fedea288	Update release notes and header to reflect recent stdio fixes.	2022-07-04 22:28:45 -05:00
Stephen Heumann	06bf0c5f46	Remove macro definition of rewind() which does not clear the IO error indicator. Now rewind() will always be called as a function. In combination with an update to the rewind() function in ORCALib, this will ensure that the error indicator is always cleared, as required by the C standards.	2022-06-24 18:32:08 -05:00
Stephen Heumann	102d6873a3	Fix type checking and result type computation for ? : operator. This was non-standard in various ways, mainly in regard to pointer types. It has been rewritten to closely follow the specification in the C standards. Several helper functions dealing with types have been introduced. They are currently only used for ? :, but they might also be useful for other purposes. New tests are also introduced to check the behavior for the ? : operator. This fixes #35 (including the initializer-specific case).	2022-06-23 22:05:34 -05:00
Stephen Heumann	802ba3b0ba	Make unary & always yield a pointer type, not an array. This affects expressions like &a (where a is an array) or &"string". In most contexts, these undergo array-to-pointer conversion anyway, but as an operand of sizeof they do not. This leads to sizeof either giving the wrong value (the size of the array rather than of a pointer) or reporting an error when the array size is not recorded as part of the type (which is currently the case for string constants). In combination with an earlier patch, this fixes #8.	2022-06-18 18:53:29 -05:00
Stephen Heumann	91b63f94d3	Note an error in the manual.	2022-06-17 18:45:59 -05:00
Stephen Heumann	67ffeac7d4	Use the proper type for expressions like &"string". These should have a pointer-to-array type, but they were treated like pointers to the first element.	2022-06-17 18:45:11 -05:00
Stephen Heumann	5e08ef01a9	Use quotes around "C" locale in release notes. This is consistent with the usage in the C standards.	2022-06-15 21:54:11 -05:00
Stephen Heumann	161bb952e3	Dynamically allocate string space, and make it larger. This increases the limit on total bytes of strings in a function, and also frees up space in the blank segment.	2022-06-08 22:09:30 -05:00
Stephen Heumann	3c2b492618	Add support for compound literals within functions. The basic approach is to generate a single expression tree containing the code for the initialization plus the reference to the compound literal (or its address). The various subexpressions are joined together with pc_bno pcodes, similar to the code generated for the comma operator. The initializer expressions are placed in a balanced binary tree, so that it is not excessively deep. Note: Common subexpression elimination has poor performance for very large trees. This is not specific to compound literals, but compound literals for relatively large arrays can run into this issue. It will eventually complete and generate a correct program, but it may be quite slow. To avoid this, turn off CSE.	2022-06-08 21:34:12 -05:00
Stephen Heumann	58771ec71c	Do not do macro expansion after each ## operator is evaluated. It should only be done after all the ## operators in the macro have been evaluated, potentially merging together several tokens via successive ## operators. Here is an example illustrating the problem: #define merge(a,b,c) a##b##c #define foobar #define foobarbaz a int merge(foo,bar,baz) = 42; int main(void) { return a; }	2022-05-24 22:38:56 -05:00
Stephen Heumann	deca73d233	Properly expand macros that have the same name as a keyword or typedef. If such macros were used within other macros, they would generally not be expanded, due to the order in which operations were evaluated during preprocessing. This is actually an issue that was fixed by the changes from ORCA/C 2.1.0 to 2.1.1 B3, but then broken again by commit `d0b4b75970`. Here is an example with the name of a keyword: #define X long int #define long X x; int main(void) { return sizeof(x); /* should be sizeof(int) / } Here is an example with the name of a typedef: typedef short T; #define T long #define X T X x; int main(void) { return sizeof(x); / should be sizeof(long) */ }	2022-05-24 22:22:37 -05:00
Stephen Heumann	a1d57c4db3	Allow ORCA/C-specific keywords to be disabled via a new pragma. This allows those tokens (asm, comp, extended, pascal, and segment) to be used as identifiers, consistent with the C standards. A new pragma (#pragma extensions) is introduced to control this. It might also be used for other things in the future.	2022-03-26 18:45:47 -05:00
Stephen Heumann	b2edeb4ad1	Properly stringize tokens that start with a trigraph. This did not work correctly before, because such tokens were recorded as starting with the third character of the trigraph. Here is an example affected by this: #define mkstr(a) # a #include <stdio.h> int main(void) { puts(mkstr(??!)); puts(mkstr(??!??!)); puts(mkstr('??<')); puts(mkstr(+??!)); puts(mkstr(+??')); }	2022-03-25 18:10:13 -05:00
Stephen Heumann	f531f38463	Use suffixes on numeric constants in #pragma expand output. A suffix will now be printed on any integer constant with a type other than int, or any floating constant with a type other than double. This ensures that all constants have the correct types, and also serves as documentation of the types.	2022-03-01 19:46:14 -06:00
Stephen Heumann	182cf66754	Properly stringize tokens with line continuations or non-initial trigraphs. Previously, continuations or trigraphs would be included in the string as-is, which should not be the case because they are (conceptually) processed in earlier compilation phases. Initial trigraphs still do not get stringized properly, because the token starting position is not recorded correctly for them. This fixes code like the following: #define mkstr(a) # a #include <stdio.h> int main(void) { puts(mkstr(a\ bc)); puts(mkstr(qr\ )); puts(mkstr(\ xy)); puts(mkstr(12??/ 34)); puts(mkstr('??<')); }	2022-03-01 19:01:11 -06:00
Stephen Heumann	fec7b57ec2	Generate a string representation of tokens merged with ##. This is necessary for correct behavior if such tokens are subsequently stringized with #. Previously, only the first half of the token would be produced. Here is an example demonstrating the issue: #define mkstr(a) # a #define in_between(a) mkstr(a) #define joinstr(a,b) in_between(a ## b) #include <stdio.h> int main(void) { puts(joinstr(123,456)); puts(joinstr(abc,def)); puts(joinstr(dou,ble)); puts(joinstr(+,=)); puts(joinstr(:,>)); }	2022-02-22 18:48:34 -06:00
Stephen Heumann	f2d6625300	Save #pragma path directives in sym files. They were not being saved, which would result in ORCA/C not searching the proper paths when looking for an include file after the sym file had ended. Here is an example showing the problem: #pragma path "include" #include <stdio.h> int k = 50; #include "n.h" /* will not find include:n.h */	2022-02-15 21:27:35 -06:00
Stephen Heumann	3893db1346	Make sure #pragma expand is properly applied in all cases. There were various places where the flag for macro expansions was saved, set to false, and then later restored. If #pragma expand was used within those areas, it would not be properly applied. Here is an example showing that problem: void f(void #pragma expand 1 ) {} This could also affect some uses of #pragma expand within precompiled headers, e.g.: #pragma expand 1 #include "a.h" #undef foobar #include "b.h" ... Also, add a note saying that code in precompiled headers will not be expanded. (This has always been the case, but was not clearly documented.)	2022-02-15 20:50:02 -06:00
Stephen Heumann	b493dcb1da	Add lint check to require whitespace after names of object-like macros. This is a requirement added in C99, so it is added as part of the C99 syntax checks. This affects definitions like: #define foo;	2022-02-13 19:44:56 -06:00
Stephen Heumann	5d7c002819	Fix bug causing some #undefs to be ignored when using a sym file. This would occur if the macro had already been saved in the sym file and the #undef occurred before a subsequent #include that was also recorded in the sym file. The solution is simply to terminate sym file generation if an #undef of an already-saved macro is encountered. Here is an example showing the problem: test.c: #include "test1.h" #undef x #include "test2.h" int main(void) { #ifdef x return x; #else return y; #endif } test1.h: #define x 27 test2.h: #define y 6	2022-02-13 16:33:43 -06:00
Stephen Heumann	b231782442	Add option to use a custom pre-include file. This is a file that will be included before the source file is processed. If specified, it is used instead of the default .h file.	2022-02-12 21:36:39 -06:00
Stephen Heumann	913a333f9f	Record the cc= string in the symbol file and require it to match. Macros and include paths from the cc= parameters may be included in the symbol file, so incorrect behavior could result if the symbol file was used for a later compilation with different cc= parameters.	2022-02-12 19:45:04 -06:00
Stephen Heumann	bd811559d6	Fix issues with keep names in sym files. There were a couple issues that could occur with #pragma keep and sym files: If a source file used #pragma keep but it was overridden by KEEP= on the command line or {KeepName} in the shell, then the overriding keep name would be saved to the sym file. It would therefore be applied to subsequent compilations even if it was no longer specified in the command line or shell variable. If a source file used #pragma keep, that keep name would be recorded in the sym file. On subsequent compilations, it would always be used, overriding any keep name specified by the command line or shell, contrary to the usual rule that the name on the command line takes priority. With this patch, the keep name recorded in the sym file (if any) should always be the one specified by #pragma keep, but it can be overridden as usual.	2022-02-06 21:49:08 -06:00
Stephen Heumann	9cdf199c3a	Clarify that sym files still need to be deleted when adding defaults.h. The old wording made it sound like it applied only to .sym files generated by an old version of ORCA/C, but that is not the case.	2022-02-06 19:06:51 -06:00
Stephen Heumann	785a6997de	Record source file changes within a function as part of debug info. This affects functions whose body spans multiple files due to includes, or is treated as doing so due to #line directives. ORCA/C will now generate a COP 6 instruction to record each source file change, allowing debuggers to properly track the flow of execution across files.	2022-02-05 18:32:11 -06:00
Stephen Heumann	7322428e1d	Add an option to print file names in error messages. This can help identify if an error is in the main source file or an include file.	2022-02-04 22:10:50 -06:00
Stephen Heumann	4cb2106ee4	Change the name of the current source file on an #include or #append. This causes __FILE__ to give the name of an include file if used within it, which seems to be what the standards intend (and what other compilers do). It also affects the file name recorded in debugging information for functions declared in an include file. (Note that occ will generate a #line directive before an #append, essentially to work around the problem this patch fixes. After the patch, such a #line directive is effectively ignored. This should be OK, although it may result in a difference in whether a full or partial pathname is used for __FILE__ and in debug info.)	2022-02-03 22:22:33 -06:00
Stephen Heumann	dce9d36edd	Comment out unused error messages and update docs about errors.	2022-02-01 22:16:57 -06:00
Stephen Heumann	e36503508a	Allow more forms of address expressions in static initializers. There were several forms that should be permitted but were not, such as &"str"[1], &"str", &a (where a is an array), and &f (where f is a function). This fixes #15 and also certain other cases illustrated in the following example: char a[10]; int main(void); static char s1 = &"string"[1]; static char s2 = &"string"; static char s3 = &a; static int (f2)(void)=&main;	2022-01-29 21:59:25 -06:00
Stephen Heumann	f4b0993007	Specify correct location for the default .h file.	2022-01-17 18:27:39 -06:00
Stephen Heumann	6f0b94bb7c	Allow the pascal qualifier to appear anywhere types are used. This is necessary to allow declarations of pascal-qualified function pointers as members of a structure, among other things. Note that the behavior for "pascal" now differs from that for the standard function specifiers, which have more restrictive rules for where they can be used. This is justified by the fact that the "pascal" qualifier is allowed and meaningful for function pointer types, so it should be able to appear anywhere they can. This fixes #28.	2022-01-13 20:11:43 -06:00
Stephen Heumann	b1bc840ec8	Reverse order of parameters for pascal function pointer types. The parameters of the underlying function type were not being reversed when applying the "pascal" qualifier to a function pointer type. This resulted in the parameters not being in the expected order when a call was made using such a function pointer. This could result in spurious errors in some cases or inappropriate parameter conversions in others. This fixes #75.	2022-01-13 19:38:22 -06:00
Stephen Heumann	3acf5844c2	Save and restore type spec when evaluating expressions in a type name. Failing to do this could allow the type spec to be overwritten if the expression contained another type name within it (e.g. a cast). This could cause the wrong type to be computed, which could lead to incorrect behavior for constructs that use type names, e.g. sizeof. Here is an example program that demonstrated the problem: int main(void) { return sizeof(short[(long)50]); }	2022-01-12 21:53:23 -06:00
Stephen Heumann	4e59f4569f	Note that structs and unions are passed by value, not by reference.	2022-01-12 18:20:21 -06:00
Stephen Heumann	de5fa5bfac	Update release notes. This adds references to some more new features to the section with manual updates.	2022-01-02 21:46:53 -06:00
Stephen Heumann	bccbcb132b	Add headers and docs for additional functions.	2021-12-24 15:57:29 -06:00

1 2 3

131 Commits