AWS 069: EBS-backed and instance-store instances
The problem
A team uses EBS-backed and instance-store instances as a label or Console setting without proving the identity, scope, behavior, failure boundary, cost, or operational result. The configuration appears complete but the real requirement remains untested.
Final outcome
The learner will inspect, explain, test, and document EBS-backed and instance-store instances. The submission must connect the Console and CLI view to the same AWS control, state what the evidence proves, diagnose one likely failure, and make a requirement-based architecture decision.
Learning outcomes
You will be able to:
- explain ebs is persistent block storage;
- explain instance store is ephemeral local storage;
- explain root device type controls stop support;
- explain ebs snapshots are incremental;
- explain performance is configured and bounded;
- distinguish successful configuration from successful workload behavior;
- preserve redacted evidence and complete the stated cleanup or retention decision.
Mental model
requirement
-> identity and permission
-> account, Region, and resource scope
-> configuration or request
-> observable state
-> workload result
-> failure evidence
-> cleanup or controlled retention
Never begin with a create button or a copied command. Start with the result that must be proven and the boundary that must remain protected.
Core facts
| Concept | What it means in practice |
|---|---|
| EBS is persistent block storage | An EBS volume exists independently in one Availability Zone, can be detached and reattached in that AZ, supports snapshots, and normally survives instance stop and start. |
| Instance store is ephemeral local storage | It is physically attached to the host. Data is lost when the instance stops, hibernates, terminates, or the underlying disk or host fails. A reboot normally preserves it. |
| Root device type controls stop support | EBS-backed instances can stop and start. Instance-store-backed instances cannot be stopped and must be replaced; support is limited to compatible AMIs and types. |
| EBS snapshots are incremental | Snapshots persist block backups in an AWS-managed regional service and can create volumes in any AZ in that Region. Copy snapshots for another Region or account. |
| Performance is configured and bounded | EBS volume type, size, IOPS, throughput, initialization, and the instance EBS bandwidth all affect performance. The lowest relevant limit becomes the bottleneck. |
| Instance store fits reproducible temporary data | Use it for caches, buffers, scratch data, and replicated systems that tolerate loss. Never keep the only copy of required data there. |
How to reason about it
Match storage to durability and latency. EBS gives persistent zonal block storage and snapshot recovery. Instance store provides fast local capacity without persistence. File and object services solve different sharing and access patterns and should not be forced into a block-storage decision.
Use the following decision table as a starting point, then change the answer when the scenario changes.
| Requirement | Preferred direction | Why |
|---|---|---|
| Boot volume and durable application files | Encrypted EBS | The data survives stop/start and supports snapshots. |
| Large rebuildable sort workspace | Instance store | Local temporary performance is useful when loss is acceptable. |
| Data must be shared by instances in several AZs | Regional file or object service | Ordinary EBS is zonal block storage, not a multi-AZ shared file system. |
| Fast recovery in another Region | Copy snapshots or replicate data before failure | A regional snapshot does not automatically exist in the recovery Region. |
Prerequisites, permissions, Region, and cost
Run this lesson as the approved non-root course identity. Begin with aws sts get-caller-identity, confirm the private account record, and set the fixed project Region before any regional query. IAM resources are global within an account, while STS endpoints and the services reached by an identity can be regional.
Use only the read or change actions required for this lesson. An AccessDenied result is evidence to analyse, not permission to switch to root or attach AdministratorAccess. Record the action, resource, principal type, request context, and smallest justified correction.
The listed cost tier is T0 - no resource creation. Free Tier eligibility and credits are account-specific. Before a mutating lab, identify every resource that can charge, estimate its duration, start a timer, and write the reverse cleanup order. A budget reports cost after billing data arrives and is not a real-time stop control.
AWS Management Console method
- Open EC2 Instances and inspect Root device type and Block devices for a learning instance.
- Open EBS Volumes to review Availability Zone, type, size, IOPS, throughput, encryption, attachments, and monitoring.
- Open the instance-type Storage specification to identify local instance-store devices and document what will be lost after stop or host failure.
For every step record the service page, selected account and Region, exact object, visible state, and why that state matters. A Console label or green status is not enough unless it is tied to the final workload outcome.
AWS CLI or API evidence
Map every attached device to EBS or instance store and inspect termination behavior.
aws ec2 describe-instances --instance-ids i-0123456789abcdef0 --query 'Reservations[0].Instances[0].{RootType:RootDeviceType,RootName:RootDeviceName,Devices:BlockDeviceMappings}'
aws ec2 describe-volumes --filters Name=attachment.instance-id,Values=i-0123456789abcdef0 --query 'Volumes[].{Volume:VolumeId,AZ:AvailabilityZone,Type:VolumeType,Size:Size,IOPS:Iops,Throughput:Throughput,Encrypted:Encrypted}' --output table
Instance-store mappings are not returned as EBS volumes. Check the instance-type specification and operating-system devices as well.
Before running the command, replace every placeholder, explain each option, and decide whether the operation is read-only or mutating. Capture the exit code immediately. Redact account IDs, ARNs, public addresses, request identifiers, and personal data before sharing.
The CLI and Console are clients of AWS APIs. Matching state across them increases confidence, but neither substitutes for data-plane or application verification.
Practical work
Create p04-storage-plan.md. Compare the 8 GiB encrypted gp3 EBS root volume with instance store. Record persistence across stop, termination deletion flag, snapshot support, attachment scope, encryption, replacement behavior, and recovery. The project requires EBS-backed root storage because the instance will stop, snapshot data, and retain controlled state.
The evidence package must include:
- non-root principal type, account verified privately, and Region;
- exact intended and observed state;
- one Console observation and the matching CLI or API result;
- one successful result and one denied, failed, or counterexample result;
- what each result does not prove;
- cost state and elapsed lab time;
- cleanup evidence or a named retained-state owner and next lesson.
Verification standard
Use three levels of proof:
- Control plane: the object or policy exists with the intended configuration.
- Data plane or behavior: the request, packet, session, storage path, or application does what the requirement states.
- Operations: monitoring, failure diagnosis, cost, ownership, and cleanup are known.
A control-plane response can precede final readiness. Use waiters or state polling where supported, then test the actual behavior. If a request times out, do not assume it failed. Inspect state and use documented idempotency before retrying a mutation.
Troubleshooting method
| Symptom | Evidence first | Smallest safe response |
|---|---|---|
| command cannot authenticate | credential source, expiry, caller preflight | restore approved temporary login |
| access is denied | action, resource, principal, all policy layers | correct only the missing or conflicting control |
| object appears missing | account, Region, filters, pagination, permission | align scope before creating anything |
| configured state exists but behavior fails | route, identity, dependency, logs, service state | test the next boundary in the path |
| cleanup is blocked | dependency inventory and owning service | remove dependants in reviewed reverse order |
Keep the original symptom and timestamp. State one hypothesis, make one reversible change, repeat the original test, and record rollback. Never open a management port to the world, expose credentials, disable TLS verification, format an unknown disk, or add broad permissions as a generic fix.
Architecture and certification decisions
Professional-level questions provide competing valid features. Identify the requirement that decides between them: human or workload identity, same-account or cross-account access, regional or zonal scope, stateful or stateless filtering, durable or ephemeral data, latency, RTO/RPO, cost, or operational ownership.
Explain why the selected option fits and why each plausible alternative fails one stated requirement. Do not rely on feature memorization or reproduce protected certification questions.
Knowledge check
- What happens to instance-store data after stop?
Expected direction: It is lost.
- Can an EBS volume attach to an instance in another AZ?
Expected direction: Not directly. Create or copy a snapshot and make a volume in the target AZ.
- Does reboot erase instance store?
Expected direction: Normally no, but the data remains non-durable and vulnerable to host loss.
- Can increasing EBS IOPS overcome a lower instance EBS limit?
Expected direction: No. The instance-side bandwidth or IOPS ceiling can remain the bottleneck.
Cost, cleanup, and retained state
Retain only the state explicitly required by the next P04 lesson and record it in the private resource ledger.
Cleanup evidence includes the final state query, not only a successful delete response. Search related resources, other Regions used by the lab, retained storage, public IPv4 addresses, logging destinations, and service-managed dependencies. Schedule a later billing review because cost data can lag.
Completion gate
Pass when the practical artifact explains the real problem, matches Console and CLI evidence, answers every knowledge check, diagnoses one failure without broadening access, records cost, and proves cleanup or approved retention. The learner must defend one decision orally and name the requirement that would change it.