Assembly Programming I CSE 351 Spring 2017

Slides:

Advertisements

Similar presentations

Introduction to Machine/Assembler Language Noah Mendelsohn Tufts University Web:

Advertisements

Machine/Assembler Language Putting It All Together Noah Mendelsohn Tufts University Web:

University of Washington Instruction Set Architectures ISAs Brief history of processors and architectures C, assembly, machine code Assembly basics: registers,

Instruction Set Architectures

COMP3221: Microprocessors and Embedded Systems Lecture 2: Instruction Set Architecture (ISA) Lecturer: Hui Wu Session.

PC hardware and x86 3/3/08 Frans Kaashoek MIT

1 ICS 51 Introductory Computer Organization Fall 2006 updated: Oct. 2, 2006.

1 Machine-Level Programming I: Basics Computer Systems Organization Andrew Case Slides adapted from Jinyang Li, Randy Bryant and Dave O’Hallaron.

6.828: PC hardware and x86 Frans Kaashoek

IT253: Computer Organization Lecture 4: Instruction Set Architecture Tonga Institute of Higher Education.

INSTRUCTION SET AND ASSEMBLY LANGUAGE PROGRAMMING

University of Washington Roadmap 1 car *c = malloc(sizeof(car)); c->miles = 100; c->gals = 17; float mpg = get_mpg(c); free(c); Car c = new Car(); c.setMiles(100);

University of Washington Basics of Machine Programming The Hardware/Software Interface CSE351 Winter 2013.

Chapter 2 Parts of a Computer System. 2.1 PC Hardware: Memory.

Carnegie Mellon 1 Machine-Level Programming I: Basics Lecture, Feb. 21, 2013 These slides are from website which accompanies the.

1 x86 Programming Model Microprocessor Computer Architectures Lab Components of any Computer System Control – logic that controls fetching/execution of.

1 Machine-Level Programming II: Basics Comp 21000: Introduction to Computer Organization & Systems Spring 2016 Instructor: John Barr * Modified slides.

Microprocessors CSE- 341 Dr. Jia Uddin Assistant Professor, CSE, BRAC University Dr. Jia Uddin, CSE, BRAC University.

Microprocessors CSE- 341 Dr. Jia Uddin Assistant Professor, CSE, BRAC University Dr. Jia Uddin, CSE, BRAC University.

Spring 2016Assembly Review Roadmap 1 car *c = malloc(sizeof(car)); c->miles = 100; c->gals = 17; float mpg = get_mpg(c); free(c); Car c = new Car(); c.setMiles(100);

Computer Science 516 Intel x86 Overview. Intel x86 Family Eight-bit 8080, 8085 – 1970s 16-bit 8086 – was internally 16 bits, externally 8 bits.

Spring 2016Machine Code & C Roadmap 1 car *c = malloc(sizeof(car)); c->miles = 100; c->gals = 17; float mpg = get_mpg(c); free(c); Car c = new Car(); c.setMiles(100);

Instruction Set Architecture

Machine-Level Programming I: Basics

Assembly Programming I CSE 351 Spring 2017

Homework Reading Lab with your assigned section starts next week

Assembly language.

Credits and Disclaimers

Assembly Programming II CSE 351 Spring 2017

Final exam: Wednesday, March 20, 2:30pm

IA32 Processors Evolutionary Design

Arrays CSE 351 Spring 2017 Instructor: Ruth Anderson

Roadmap C: Java: Assembly language: OS: Machine code: Computer system:

Assembly Programming I CSE 410 Winter 2017

Roadmap C: Java: Assembly language: OS: Machine code: Computer system:

Computer Organization & Assembly Language Chapter 3

Today How’s Lab 3 going? HW 3 will be out today

The Stack & Procedures CSE 351 Spring 2017

Basics Of X86 Architecture

x86 Programming I CSE 351 Autumn 2016

Homework Reading Continue work on mp1

Introduction to Assembly Language

BIC 10503: COMPUTER ARCHITECTURE

Assembly Programming IV CSE 351 Spring 2017

x86-64 Programming I CSE 351 Summer 2018

x86-64 Assembly CSE 351 Winter 2018 Instructor: Mark Wyse

Introduction to Intel IA-32 and IA-64 Instruction Set Architectures

Lecture 30 (based on slides by R. Bodik)

The University of Adelaide, School of Computer Science

x86-64 Programming I CSE 351 Autumn 2017

Symbolic Instruction and Addressing

x86-64 Assembly CSE 351 Winter 2018 Instructor: Mark Wyse

Roadmap C: Java: Assembly language: OS: Machine code: Computer system:

Computer Architecture CST 250

x86-64 Programming I CSE 351 Winter 2018

Introduction to Microprocessor Programming

x86-64 Programming I CSE 351 Winter 2018

What is Computer Architecture?

x86-64 Programming I CSE 351 Autumn 2018

Machine-Level Programming II: Basics Comp 21000: Introduction to Computer Organization & Systems Instructor: John Barr * Modified slides from the book.

Machine-Level Programming I: Basics

Machine-Level Programming II: Basics Comp 21000: Introduction to Computer Organization & Systems Spring 2016 Instructor: John Barr * Modified slides.

Floating Point II, x86-64 Intro CSE 351 Autumn 2017

CSC 497/583 Advanced Topics in Computer Security

Machine-Level Programming I: Basics Comp 21000: Introduction to Computer Organization & Systems Instructor: John Barr * Modified slides from the book.

The von Neumann Machine

Credits and Disclaimers

x86-64 Programming I CSE 351 Spring 2019

Low level Programming.

Presentation transcript:

Assembly Programming I CSE 351 Spring 2017 Instructor: Ruth Anderson Teaching Assistants: Dylan Johnson Kevin Bi Linxing Preston Jiang Cody Ohlsen Yufang Sun Joshua Curtis

Administrivia Lab 1 due Friday (4/14) Prelim submission (3+ of bits.c) due on TONIGHT (4/10). Turn in whatever you have at that time (drop box closes at 11:59pm), no lates. Worth a small part (no more than 10%) of total points for lab 1. Homework 2 due next Wednesday (4/19)

Roadmap C: Java: Assembly language: OS: Machine code: Computer system: Memory & data Integers & floats x86 assembly Procedures & stacks Executables Arrays & structs Memory & caches Processes Virtual memory Memory allocation Java vs. C C: Java: car *c = malloc(sizeof(car)); c->miles = 100; c->gals = 17; float mpg = get_mpg(c); free(c); Car c = new Car(); c.setMiles(100); c.setGals(17); float mpg = c.getMPG(); Assembly language: get_mpg: pushq %rbp movq %rsp, %rbp ... popq %rbp ret OS: Machine code: 0111010000011000 100011010000010000000010 1000100111000010 110000011111101000011111 Computer system:

What makes programs run fast(er)? Translation Code Time Compile Time Run Time User program in C Hardware C compiler Assembler .c file .exe file What makes programs run fast(er)?

HW Interface Affects Performance Source code Compiler Architecture Hardware Instruction set Different implementations Different applications or algorithms Perform optimizations, generate instructions Intel Pentium 4 C Language Program A Intel Core 2 GCC x86-64 Intel Core i7 Program B AMD Opteron Clang AMD Athlon Different programs written in C, have different algorithms Compilers figure out what instructions to generate do different optimizations Different architectures expose different interfaces Hardware implementations have different performance Your program ARMv8 (AArch64/A64) ARM Cortex-A53 Apple A7

Instruction Set Architectures The ISA defines: The system’s state (e.g. registers, memory, program counter) The instructions the CPU can execute The effect that each of these instructions will have on the system state CPU Memory PC Registers Textbook: “the format and behavior of a machine-level program is defined by the ISA, defining the processor state, the format of the instructions, and the effect each of these instructions will have on the state.”

Instruction Set Philosophies Complex Instruction Set Computing (CISC): Add more and more elaborate and specialized instructions as needed Lots of tools for programmers to use, but hardware must be able to handle all instructions x86-64 is CISC, but only a small subset of instructions encountered with Linux programs Reduced Instruction Set Computing (RISC): Keep instruction set small and regular Easier to build fast hardware Let software do the complicated operations by composing simpler ones

General ISA Design Decisions Instructions What instructions are available? What do they do? How are they encoded? Registers How many registers are there? How wide are they? Memory How do you specify a memory location? Registers: word size

General ISA Design Decisions Instructions What instructions are available? What do they do? How are they encoded? Instructions are data! Registers How many registers are there? How wide are they? Size of a word Memory How do you specify a memory location? Different ways to build up an address Registers: word size

Mainstream ISAs Macbooks & PCs (Core i3, i5, i7, M) x86-64 Instruction Set Smartphone-like devices (iPhone, iPad, Raspberry Pi) ARM Instruction Set Digital home & networking equipment (Blu-ray, PlayStation 2) MIPS Instruction Set x86 backwards-compatible all the way to 8086 architecture (introduced in 1978!), has had many, many new features added since. x86 dominates the server, desktop, and laptop markets. ARM stands for Advanced RISC Machine. MIPS was originally an acronym for Microprocessor without Interlocked Pipeline Stages. Developed by MIPS Technologies (formerly MIPS Computer Systems, Inc).

Definitions Architecture (ISA): The parts of a processor design that one needs to understand to write assembly code “What is directly visible to software” Microarchitecture: Implementation of the architecture CSE/EE 469, 470 Are the following part of the architecture? Number of registers? How about CPU frequency? Cache size? Memory size? Registers? Yes. Each has a unique name that the assembly programmer must know in order to use. Frequency? Nope, same program, just faster! Is size of memory part of architecture? No! Don’t’ have to change your program if you add more memory

Definitions Architecture (ISA): The parts of a processor design that one needs to understand to write assembly code “What is directly visible to software” Microarchitecture: Implementation of the architecture CSE/EE 469, 470 Are the following part of the architecture? Number of registers? Yes How about CPU frequency? No Cache size? Memory size? No Registers? Yes. Each has a unique name that the assembly programmer must know in order to use. Frequency? Nope, same program, just faster! Is size of memory part of architecture? No! Don’t’ have to change your program if you add more memory

Assembly Programmer’s View CPU Memory Addresses Registers PC Code Data Stack Data Condition Codes Instructions Programmer-visible state PC: the Program Counter (%rip in x86-64) Address of next instruction Named registers Together in “register file” Heavily used program data Condition codes Store status information about most recent arithmetic operation Used for conditional branching Memory Byte-addressable array Code and user data Includes the Stack (for supporting procedures) Preview for the coming lectures

x86-64 Assembly “Data Types” Integral data of 1, 2, 4, or 8 bytes Data values Addresses (untyped pointers) Floating point data of 4, 8, 10 or 2x8 or 4x4 or 8x2 Different registers for those (e.g. %xmm1, %ymm2) Come from extensions to x86 (SSE, AVX, …) No aggregate types such as arrays or structures Just contiguously allocated bytes in memory Two common syntaxes “AT&T”: used by our course, slides, textbook, gnu tools, … “Intel”: used by Intel documentation, Intel tools, … Must know which you’re reading Not covered In CSE 351 Assembly program doesn’t actually have notion of “data types,” but since encoding is so different between integral and floating point data, there are actually different instructions built into assembly because the hardware needs to deal with these numbers differently when performing arithmetic.

What is a Register? A location in the CPU that stores a small amount of data, which can be accessed very quickly (once every clock cycle) Registers have names, not addresses In assembly, they start with % (e.g. %rsi) Registers are at the heart of assembly programming They are a precious commodity in all architectures, but especially x86

x86-64 Integer Registers – 64 bits wide %rsp %esp %eax %rax %ebx %rbx %ecx %rcx %edx %rdx %esi %rsi %edi %rdi %ebp %rbp %r8d %r8 %r9d %r9 %r10d %r10 %r11d %r11 %r12d %r12 %r13d %r13 %r14d %r14 %r15d %r15 Can reference low-order 4 bytes (also low-order 2 & 1 bytes) Names are largely arbitrary, reasons behind them are not super relevant to us But there are conventions about how they are used One is particularly important to not misuse: %rsp Stack pointer; we’ll get to this when we talk about procedures Why is %rbp now general-purpose in x86-64? http://www.codemachine.com/article_x64deepdive.html r8-r15: still names not indexes Each is 8-bytes wide What if we only want 4 bytes? We can use these alternate names to use just half of the bytes

Some History: IA32 Registers – 32 bits wide %esi %si %edi %di %esp %sp %ebp %bp %eax %ax %ah %al %ecx %cx %ch %cl %edx %dx %dh %dl %ebx %bx %bh %bl 16-bit virtual registers (backwards compatibility) general purpose accumulate counter data base source index destination index stack pointer base pointer Name Origin (mostly obsolete) Only 6 registers that we can do whatever we want with Remember how I said we’ve been upgrading x86 since 1978, when it was 16-bit? Originally just %ax, … Could even reference single bytes (high, low) There used to be more conventions about how they were used Just to make it easier when people were writing assembly by hand Now irrelevant

Memory vs. Registers Addresses vs. Names Big vs. Small Slow vs. Fast 0x7FFFD024C3DC %rdi Big vs. Small ~ 8 GiB (16 x 8 B) = 128 B Slow vs. Fast ~50-100 ns sub-nanosecond timescale Dynamic vs. Static Can “grow” as needed fixed number in hardware while program runs

Three Basic Kinds of Instructions Transfer data between memory and register Load data from memory into register %reg = Mem[address] Store register data into memory Mem[address] = %reg Perform arithmetic operation on register or memory data c = a + b; z = x << y; i = h & g; Control flow: what instruction to execute next Unconditional jumps to/from procedures Conditional branches Remember: Memory is indexed just like an array of bytes!

Operand types Immediate: Constant integer data %rax %rcx Examples: $0x400, $-533 Like C literal, but prefixed with ‘$’ Encoded with 1, 2, 4, or 8 bytes depending on the instruction Register: 1 of 16 integer registers Examples: %rax, %r13 But %rsp reserved for special use Others have special uses for particular instructions Memory: Consecutive bytes of memory at a computed address Simplest example: (%rax) Various other “address modes” %rax %rcx %rdx %rbx %rsi %rdi %rsp %rbp %rN

Moving Data General form: mov_ source, destination movb src, dst Missing letter (_) specifies size of operands Note that due to backwards-compatible support for 8086 programs (16-bit machines!), “word” means 16 bits = 2 bytes in x86 instruction names Lots of these in typical code movb src, dst Move 1-byte “byte” movw src, dst Move 2-byte “word” movl src, dst Move 4-byte “long word” movq src, dst Move 8-byte “quad word”

movq Operand Combinations Source Dest Src, Dest C Analog movq Imm Reg movq $0x4, %rax Mem movq $-147, (%rax) movq %rax, %rdx movq %rax, (%rdx) movq (%rax), %rdx var_a = 0x4; *p_a = -147; var_d = var_a; *p_d = var_a; var_d = *p_a; Cannot do memory-memory transfer with a single instruction How would you do it?

Question Which of the following statements is TRUE? For float f, (f+2 == f+1+1) always returns TRUE The width of a “word” is part of a system’s architecture (as opposed to microarchitecture) Having more registers increases the performance of the hardware, but decreases the performance of the software Mem to Mem (src to dst) is the only disallowed operand combination in x86-64 B is true -D sounds true, but keep in mind that you cannot write to a immediate value

Summary Converting between integral and floating point data types does change the bits Floating point rounding is a HUGE issue! Limited mantissa bits cause inaccurate representations Floating point arithmetic is NOT associative or distributive x86-64 is a complex instruction set computing (CISC) architecture Registers are named locations in the CPU for holding and manipulating data x86-64 uses 16 64-bit wide registers Assembly operands include immediates, registers, and data at specified memory locations