Chapter 3: Scanning Source Code

Book Overview & Buying
Table Of Contents

Build Your Own Programming Language

By : Clinton L. Jeffery

4.4 (17)

Buy this Book

Build Your Own Programming Language

4.4 (17)

By: Clinton L. Jeffery

Buy this Book

Overview of this book

The need for different types of computer languages is growing rapidly and developers prefer creating domain-specific languages for solving specific application domain problems. Building your own programming language has its advantages. It can be your antidote to the ever-increasing size and complexity of software. In this book, you’ll start with implementing the frontend of a compiler for your language, including a lexical analyzer and parser. The book covers a series of traversals of syntax trees, culminating with code generation for a bytecode virtual machine. Moving ahead, you’ll learn how domain-specific language features are often best represented by operators and functions that are built into the language, rather than library functions. We’ll conclude with how to implement garbage collection, including reference counting and mark-and-sweep garbage collection. Throughout the book, Dr. Jeffery weaves in his experience of building the Unicon programming language to give better context to the concepts where relevant examples are provided in both Unicon and Java so that you can follow the code of your choice of either a very high-level language with advanced features, or a mainstream language. By the end of this book, you’ll be able to build and deploy your own domain-specific languages, capable of compiling and running programs.

Preface

Who this book is for

What this book covers

To get the most out of this book

Download the example code files

Code in Action

Download the color images

Conventions used

Get in touch

Share Your Thoughts

Section 1: Programming Language Frontends

Free Chapter

Chapter 1: Why Build Another Programming Language?

So, you want to write your own programming language…

Language versus library – what's the difference?

Applicability to other software engineering tasks

Establishing the requirements for your language

Case study – requirements that inspired the Unicon language

Summary

Questions

Chapter 2: Programming Language Design

Determining the kinds of words and punctuation to provide in your language

Specifying the control flow

Deciding on what kinds of data to support

Overall program structure

Completing the Jzero language definition

Case study – designing graphics facilities in Unicon

Summary

Questions

Chapter 3: Scanning Source Code

Technical requirements

Lexemes, lexical categories, and tokens

Regular expressions

Using UFlex and JFlex

Writing a scanner for Jzero

Regular expressions are not always enough

Summary

Questions

Chapter 4: Parsing

Technical requirements

Analyzing syntax

Understanding context-free grammars

Using iyacc and BYACC/J

Writing a parser for Jzero

Improving syntax error messages

Summary

Questions

Chapter 5: Syntax Trees

Technical requirements

Using GNU make

Learning about trees

Creating leaves from terminal symbols

Building internal nodes from production rules

Forming syntax trees for the Jzero language

Debugging and testing your syntax tree

Summary

Questions

Section 2: Syntax Tree Traversals

Chapter 6: Symbol Tables

Technical requirements

Establishing the groundwork for symbol tables

Creating and populating symbol tables for each scope

Checking for undeclared variables

Finding redeclared variables

Handling package and class scopes in Unicon

Testing and debugging symbol tables

Summary

Questions

Chapter 7: Checking Base Types

Technical requirements

Type representation in the compiler

Assigning type information to declared variables

Determining the type at each syntax tree node

Runtime type checks and type inference in Unicon

Summary

Questions

Chapter 8: Checking Types on Arrays, Method Calls, and Structure Accesses

Technical requirements

Checking operations on array types

Checking method calls

Checking structured type accesses

Summary

Questions

Chapter 9: Intermediate Code Generation

Technical requirements

Preparing to generate code

An intermediate code instruction set

Annotating syntax trees with labels for control flow

Generating code for expressions

Generating code for control flow

Summary

Chapter 10: Syntax Coloring in an IDE

Downloading the example IDEs used in this chapter

Integrating a compiler into a programmer's editor

Avoiding reparsing the entire file on every change

Using lexical information to colorize tokens

Highlighting errors using parse results

Adding Java support

Summary

Section 3: Code Generation and Runtime Systems

Chapter 11: Bytecode Interpreters

Technical requirements

Understanding what bytecode is

Comparing bytecode with intermediate code

Building a bytecode instruction set for Jzero

Implementing a bytecode interpreter

Writing a runtime system for Jzero

Running a Jzero program

Examining iconx, the Unicon bytecode interpreter

Summary

Questions

Chapter 12: Generating Bytecode

Technical requirements

Converting intermediate code to Jzero bytecode

Comparing bytecode assembler with binary formats

Linking, loading, and including the runtime system

Unicon example – bytecode generation in icont

Summary

Questions

Chapter 13: Native Code Generation

Technical requirements

Deciding whether to generate native code

Introducing the x64 instruction set

Using registers

Converting intermediate code to x64 code

Generating x64 output

Summary

Questions

Chapter 14: Implementing Operators and Built-In Functions

Implementing operators

Writing built-in functions

Integrating built-ins with control structures

Developing operators and functions for Unicon

Summary

Questions

Chapter 15: Domain Control Structures

Knowing when you need a new control structure

Scanning strings in Icon and Unicon

Rendering regions in Unicon

Summary

Questions

Chapter 16: Garbage Collection

Appreciating the importance of garbage collection

Counting references to objects

Marking live data and sweeping the rest

Summary

Questions

Chapter 17: Final Thoughts

Reflecting on what was learned from writing this book

Deciding where to go from here

Exploring references for further reading

Summary

Section 4: Appendix

Assessments

Chapter 1

Chapter 2

Chapter 3

Chapter 4

Chapter 5

Chapter 6

Chapter 7

Chapter 8

Chapter 11

Chapter 12

Chapter 13

Chapter 14

Chapter 15

Chapter 16

Why subscribe?

Other Books You May Enjoy

Packt is searching for authors like you

Share Your Thoughts

Appendix: Unicon Essentials

Running Unicon

Using Unicon's declarations and data types

Evaluating expressions

Debugging and environmental issues

Function mini-reference

Selected keywords

Build Your Own Programming Language

By : Clinton L. Jeffery

Build Your Own Programming Language

By: Clinton L. Jeffery

Overview of this book

Confirmation

Buy this book with your credits?

Submit Your Feedback

Create a Free Account To Continue Reading

Sign in to activate your 7-day free access