File Operations in Shell¶
Overview¶
Backup jobs, lock files, and config rewrites all touch the filesystem. Do it with tests, temps, and cleanup traps.
This is Tutorial 9 in Module 9: File Operations of the REBASH Academy Shell Scripting for DevOps Engineers series — written for Linux administrators, DevOps engineers, SREs, and platform engineers who automate production hosts with Bash.
Prerequisites¶
- Arrays and String Manipulation
- Bash 4.2+ on Linux (WSL2/VM/cloud)
Learning Objectives¶
By the end of this tutorial, you will be able to:
- Apply the core ideas of “File Operations in Shell” in a real ops script
- Use
set -euo pipefailas the production default - Use quoted expansions and clear stderr diagnostics
- Produce meaningful exit codes for automation consumers
- Debug behaviour with
bash -xwhen something fails - Relate this topic to day-to-day Linux admin and DevOps work
Architecture¶
Ops scripts sit between humans/automation and system tools. This topic’s control points are shown below.
Theory¶
What it is¶
Shell scripts spend much of their life reading and writing files: configs, inventories, logs, lock files, and temporary workspaces. Core skills include line-oriented reading, safe writing and atomic replace, creating temporary files with cleanup traps, file tests (-f, -d, and friends), and directory operations such as mkdir -p, find, and rsync. Done well, file handling is boring and reliable; done poorly, it races, truncates, or deletes the wrong tree.
Why it matters¶
Automation that mutates the filesystem is where Linux administration meets risk. Partial writes leave corrupt configs that services load on the next restart. Fixed paths under /tmp collide between users or invite symlink attacks. Recursive deletes without an allow-listed root are a classic outage class. Continuous Integration (CI) and cron jobs that create debris without trap cleanup fill disks over time. Safe file patterns are therefore as important as the business logic that uses those files.
How it works¶
Read lines without word-splitting:
Prefer mapfile / readarray for modest files that fit in memory. Avoid for line in $(cat file). Writing uses printf ... >"$out" to truncate or >> to append. For config updates, write to a temporary file on the same filesystem, then mv over the target so readers see either the old or the new file — not a half-written hybrid.
Create temps with mktemp and always register cleanup:
Never hard-code /tmp/myjob. Before destructive directory work, test paths (-e, -f, -d, -L, permission bits) and keep operations under an allowed root. Use mkdir -p, install -d, find, and rsync -a as appropriate for trees.
Key concepts¶
| Concern | Practice |
|---|---|
| Reading | IFS= read -r or mapfile; never for $(cat) |
| Writing | Redirect carefully; atomic tmp + mv for replace |
| Temps | mktemp / mktemp -d + trap on EXIT |
| Tests | -e -f -d -r -w -x -s -L |
| Trees | mkdir -p, find, rsync; allow-list before rm -rf |
Common pitfalls¶
- Truncating a live config with
>instead of writing via a temp file - Hard-coding predictable temp paths that other users can race
- Forgetting
trapcleanup so failed runs leave large directories behind - Running
rm -rfon a variable that expanded empty or to/ - Checking only byte free space and ignoring inode exhaustion on the same volume
Hands-on Lab¶
Create a workspace for this tutorial.
Focus: read/write atomically; mktemp + trap; file tests; safe mkdir
Step 1 – Safe file ops¶
cat > files.sh << 'EOF'
#!/usr/bin/env bash
set -euo pipefail
root=$PWD/data
mkdir -p "$root"
tmp=$(mktemp)
trap 'rm -f "$tmp"' EXIT
printf 'line1\nline2\n' >"$tmp"
out=$root/out.txt
cp "$tmp" "$out.tmp"
mv "$out.tmp" "$out"
[[ -f "$out" && -s "$out" ]] || exit 4
while IFS= read -r line; do echo "got=$line"; done <"$out"
EOF
chmod +x files.sh
./files.sh
Final step – Cleanup note¶
Validation¶
- Lab commands run under
~/rebash-shell/lab09/ - You can explain each Theory heading in your own words
- Failure path exits non-zero and prints diagnostics to stderr (where applicable)
- You can relate this topic to a real DevOps or Linux admin task
Code Walkthrough¶
Production Bash for File Operations in Shell always combines:
- A clear shebang (
#!/usr/bin/env bash) - Strict mode near the top (
set -euo pipefail) from Module 2 onward - Quoted expansions and explicit tests
- Functions with
localfor reusable behaviour - Documented exit codes and stderr logging
Keep scripts short enough to review in a single merge request. When logic grows (complex JSON APIs, heavy state), hand off to Python and keep Bash as the launcher.
Security Considerations¶
- Treat all external input (args, files, env) as untrusted until validated
- Never log secrets; prefer masked CI variables and secret stores
- Prefer least privilege — do not require root for file-local tasks
- Avoid
evaland unquoted expansions in destructive commands - Validate paths stay under an allow-listed root before
rmor overwrite
Common Mistakes¶
Skipping strict mode
Cron and CI hide failures that an interactive terminal would show. Fix: start with set -euo pipefail from Module 2 onward.
Unquoted path expansions
Spaces and globs rewrite your command line. Fix: always "$path" / "$@".
Assuming interactive PATH
Aliases and fancy PATH entries disappear under schedulers. Fix: set PATH or use absolute paths.
Best Practices¶
- One purpose per script; compose with functions or small binaries
- Log to stderr; reserve stdout for data or RESULT lines
- Idempotent behaviour where scheduling may overlap
- Pair every new script with a failing-path test you actually run
- Run ShellCheck in CI before merging automation
Troubleshooting¶
| Symptom | Likely cause | Fix |
|---|---|---|
| Works in terminal, fails in cron | PATH / cwd / env | Fingerprint env; set PATH |
unbound variable | set -u | Provide defaults or export vars |
| Pipeline “succeeds” incorrectly | Missing pipefail | set -o pipefail |
[[ unexpected operator | Running under sh/dash | Fix shebang to Bash |
Summary¶
File Operations in Shell is a core skill for Linux admins and DevOps engineers automating real hosts and pipelines. Practise the lab until the failure path is as familiar as the happy path, then continue the track.
Interview Questions¶
- How does this topic show up in production Linux administration or CI?
- What failure mode appears if you ignore quoting or strict mode here?
- How would you test this behaviour under a minimal cron-like environment?
- When would you move this logic out of Bash into Python or another tool?
- What exit code contract would you document for teammates?
Sample answer — question 2
Unquoted expansions and missing pipefail create silent or partial failures — especially under cron — that look healthy in monitoring until data is wrong.
Related Tutorials¶
- Shell Scripting for DevOps Engineers – Category Overview
- Arrays and String Manipulation (previous)
- Text Processing in Shell Scripts (next)
- Learning Paths