Assembly Language & x86 Low-Level Systems Programming · 강의

데이터 표현과 타입

정수, 문자와 기타 데이터 타입이 메모리에 저장되고 어셈블리 명령어로 조작되는 방식을 배웁니다.

레슨 3/411개 단계

데이터 표현과 타입은(는) CoddyKit의 무료 Assembly Language & x86 Low-Level Systems Programming 강의입니다. 이것은 4개 중 3번째 강의입니다. 아래에서 전체 강의를 무료로 읽을 수 있으며, 내장 코드 에디터와 24/7 AI 튜터와 함께 브라우저에서 직접 실습할 수 있습니다. 이 강의는 Assembly Language & x86 Low-Level Systems Programming 학습 경로의 일부이며, 진행 상황이 웹과 CoddyKit 앱에 동기화됩니다. Assembly Language & x86 Low-Level Systems Programming 강의에는 총 4개의 강의가 포함되어 있습니다.

이 강의의 일부는 아직 번역되지 않았으며 영어로 표시됩니다.

What are Data Types?

In assembly language, we work directly with raw bits and bytes. But how do we know if a sequence of bytes represents a number, a character, or something else?

This is where data types come in! They give meaning to the raw data, helping both you and the CPU understand how to interpret and manipulate information.

Common Data Sizes

x86 assembly defines standard sizes for data. These directly correspond to how much memory space a piece of data occupies:

  • BYTE: 8 bits
  • WORD: 16 bits (2 bytes)
  • DWORD: Double Word, 32 bits (4 bytes)
  • QWORD: Quad Word, 64 bits (8 bytes)

These sizes are fundamental for declaring variables and working with registers.

Defining Data: DB, DW, DD, DQ

To store data in memory, we use data definition directives. These tell the assembler to reserve space and optionally initialize it with a value.

  • DB: Define Byte (8-bit)
  • DW: Define Word (16-bit)
  • DD: Define Doubleword (32-bit)
  • DQ: Define Quadword (64-bit)

You'll see these often when creating variables in your programs.

Unsigned Integers

An unsigned integer is a number that is always positive or zero. All of its bits are used to represent the magnitude of the number.

For example, an 8-bit unsigned byte can hold values from 0 to 255. A 16-bit unsigned word can hold values from 0 to 65,535.

When you don't need negative numbers, unsigned types are perfect and give you a larger positive range.

Signed Integers (Two's Complement)

Signed integers can represent both positive and negative values. One bit, usually the Most Significant Bit (MSB), is used to indicate the sign (0 for positive, 1 for negative).

Negative numbers are typically represented using Two's Complement. This system makes arithmetic operations work seamlessly for both positive and negative values.

An 8-bit signed byte ranges from -128 to +127.

Character Data: ASCII

Characters like 'A', 'b', or '7' are also stored as numbers! The most common standard for this is ASCII (American Standard Code for Information Interchange).

Each character is assigned a unique 8-bit (1-byte) numerical value. For example, the character 'A' is represented by the decimal value 65 (or hexadecimal 0x41).

You can define single characters or entire strings using the DB directive.

Code: Defining & Accessing Data

This example shows how to define different data types and then load their values into CPU registers. This demonstrates how assembly treats these named memory locations.

section .data
    ; Define various data types
    myByte  db 10        ; An 8-bit unsigned integer
    myWord  dw 256       ; A 16-bit unsigned integer
    myDword dd 65536     ; A 32-bit unsigned integer
    myChar  db 'X'       ; An 8-bit character (ASCII value 88)
    myString db "Hello", 0 ; A string (null-terminated)

section .text
    global _start

_start:
    ; Move byte into AL register
    mov al, [myByte]

    ; Move word into BX register
    mov bx, [myWord]

    ; Move dword into ECX register
    mov ecx, [myDword]

    ; Move char into DL register
    mov dl, [myChar]

    ; Exit gracefully (Linux syscall)
    mov eax, 1           ; sys_exit syscall number
    xor ebx, ebx         ; Exit code 0
    int 0x80             ; Invoke kernel

Data Alignment Benefits

Data alignment means placing data in memory at an address that is a multiple of its size. For example, a DWORD (4 bytes) might be aligned to an address ending in 0, 4, 8, or C (hex).

While not strictly required by all CPUs, proper alignment can significantly improve performance. The CPU can fetch aligned data more efficiently, often in a single memory access, avoiding extra work.

Assemblers sometimes provide directives like ALIGN to help ensure proper alignment.

Why Data Types Matter

Understanding data types is crucial because it dictates:

  • Memory Usage: How much space your data consumes.
  • Instruction Choice: Which assembly instructions (e.g., ADD, MOV) are appropriate for the data size.
  • Interpretation: Whether the CPU treats 0xFF as 255 (unsigned) or -1 (signed).

Careful type selection prevents errors and ensures your programs behave as expected at the lowest level.

Quick Check: Data Sizes

You've learned about common data sizes and how they're defined. Let's test your knowledge!

Recap: Data Representation

Great job! You've explored the fundamental concepts of data representation in x86 assembly.

  • We define data using directives like DB, DW, DD, and DQ for various sizes.
  • Integers can be signed (positive/negative) or unsigned (positive only).
  • Characters are stored using the ASCII standard, where each character has a numerical value.
  • Understanding data alignment can help optimize performance.

Next, we'll continue building on this knowledge to perform more complex operations!

무료로 시작

AI 튜터와 함께 Assembly을(를) 배우세요 — 무료

브라우저에서 실제 코드를 작성하고 실행하며, 24/7 AI 튜터로부터 즉각적인 도움을 받고, 웹이나 앱에서 중단한 부분부터 계속 학습하세요.

코스
12
레슨
48

자주 묻는 질문

“데이터 표현과 타입” 강의는 무료인가요?

네 — “데이터 표현과 타입” 전체 내용을 이 웹사이트에서 무료로 읽을 수 있습니다. 인터랙티브하게 실습하려면(내장 코드 에디터와 24/7 AI 튜터), CoddyKit PRO로 업그레이드하면 Assembly Language & x86 Low-Level Systems Programming 강의 전체를 잠금 해제할 수 있습니다. Assembly Language & x86 Low-Level Systems Programming 강의에는 총 4개의 강의가 포함되어 있습니다.

“데이터 표현과 타입”에서 뭘 배우나요?

정수, 문자와 기타 데이터 타입이 메모리에 저장되고 어셈블리 명령어로 조작되는 방식을 배웁니다. 브라우저에서 직접 실행하는 실습 코드로 Assembly Language & x86 Low-Level Systems Programming을(를) 배우며, 24/7 AI 튜터가 강의를 진행하면서 질문에 답변해줍니다.

Assembly Language & x86 Low-Level Systems Programming을(를) 시작하는 데 경험이 필요한가요?

사전 경험은 필요하지 않습니다. CoddyKit의 Assembly Language & x86 Low-Level Systems Programming은(는) 초급자부터 고급 학습자까지를 위해 구성되어 있으므로, 여기서 시작하거나 처음부터 시작할 수 있으며 자신의 속도대로 진행할 수 있습니다. 이것은 4개 중 3번째 강의입니다.

“데이터 표현과 타입” 강의는 얼마나 걸리나요?

대부분의 CoddyKit 강의는 약 5~10분이 소요됩니다. 각 강의는 간결하고 인터랙티브하여 꾸준한 진행이 가능하며, 웹과 앱에서 중단한 부분부터 바로 시작할 수 있습니다.

이 Assembly Language & x86 Low-Level Systems Programming 강의에서 코드를 작성하고 실행할 수 있나요?

네. 모든 Assembly Language & x86 Low-Level Systems Programming 강의에는 내장 코드 에디터가 포함되어 있으므로, 브라우저에서 바로 실제 코드를 작성하고 실행한 후 즉시 AI 피드백을 받을 수 있습니다 — 로컬 설정이 필요 없습니다.

이 강의의 모든 강의

  1. x86 레지스터 쉽게 이해하기
  2. 메모리 주소 지정 모드
  3. 데이터 표현과 타입
  4. FLAGS 레지스터와 상태 비트
← Assembly Language & x86 Low-Level Systems Programming(으)로 돌아가기