Lustre Script Coding Style: Difference between revisions
Jump to navigation
Jump to search
m (→Subtests: minor fixups) |
(→Subtests: describe conventions for subtest naming) |
||
| Line 107: | Line 107: | ||
=== Subtests === | === Subtests === | ||
* subtest names/numbers are not strictly meaningful in themselves, but there are some conventions that should be followed | |||
** prefer subtest numbers below 1000 by convention | |||
*** some wrapper scripts (e.g. sanity-compr.sh) run other scripts like sanity.sh and sanityn.sh as well as their own subtests | |||
** it is convenient to cluster related tests together when adding new subtests | |||
*** easier to find related subtests to determine if there is coverage for some functionality | |||
*** adding subtests in different parts of the script (instead of at the end) avoids needless context patch conflicts | |||
*** adding subtests with disjoint numbers avoids semantic test numbering conflicts | |||
** when adding subtests at the end of the file it is convenient to leave occasional gaps in the numbering (e.g. 100, 150, 200 instead of 33, 34, 35,) | |||
*** this leaves room for other subtests to be added with fewer context/number conflicts | |||
* '''do not''' mix pure numbered subtests with suffixes for the same number (e.g. 56 and 56a) | |||
** test-framework.sh allows running groups of subtests together (e.g. all 56X subtests), but if there is also plain-numbered subtest (e.g. 56) then it cannot be run in isolation | |||
** if a numbered subtest is growing a new variant, rename the original subtest to the <code>a</code> suffix (e.g. 56->56a) then add the new variant with the <code>b</code> suffix | |||
** if the plain a-z suffixes are occupied or you want to make a variant of an already-suffixed subtest, use <code>aa</code>, <code>ab</code> (or <code>ea</code> or similar) and locate it after <code>a</code> (or <code>e</code>) | |||
** continue a-z suffixes with A-Z, but do not add aa, ab, etc. '''after''' the z or Z suffix to avoid clashes of subtest names | |||
* subtests should not reference their test number explicitly, since the subtest name may change in the future (for various reasons). | |||
** the <code>$testnum</code> variable holds the test number (e.g. <code>11</code> or <code>27M</code>) | |||
** the <code>$TESTNAME</code> variable holds the test name (e.g. <code>test_27</code>) | |||
* test files should be named <code>$tfile</code> for the filename, or base name like <code>$tfile.1</code> or <code>$tfile.source</code> to simplify debugging | * test files should be named <code>$tfile</code> for the filename, or base name like <code>$tfile.1</code> or <code>$tfile.source</code> to simplify debugging | ||
* test directories should be named <code>$tdir</code>, and should be used insetad of <code>$DIR</code> if a | * test directories should be named <code>$tdir</code>, and should be used insetad of <code>$DIR</code> if more than a handful files are created for the subtest | ||
* small/few test files/dirs do not need to be explicitly deleted at the end of the test, that is done by test-framework.sh at the start/end of each test script | * small/few test files/dirs do not need to be explicitly deleted at the end of the test, that is done by test-framework.sh at the start/end of each test script | ||
* large (over 1MB)/many (over 50) test files/dirs in a subtest should be cleaned up explicitly with a <code>stack_trap</code> so that they are always cleaned up even if the test exits with an error, and do not fill the test filesystem | * large (over 1MB)/many (over 50) test files/dirs in a subtest should be cleaned up explicitly with a <code>stack_trap</code> so that they are always cleaned up even if the test exits with an error, and do not fill the test filesystem | ||
| Line 138: | Line 156: | ||
kinit || skip_env "Kerberos not installed" | kinit || skip_env "Kerberos not installed" | ||
</nowiki> | </nowiki> | ||
* when subtests are run in a loop with <code>ONLY_REPEAT=N</code> or <code>ONLY_MINUTES=M</code> then <code>$ONLY_REPEAT_ITER</code> holds the iteration number (1-based). In some cases this can be useful for adapting the test case to handle running many iterations in a row (e.g. adding <code>wait_delete_completed</code>) that isn't needed when run in isolation. | |||
[[Category: Development]] | [[Category: Development]] | ||
Revision as of 14:34, 14 August 2026
Bash Style
- Bash is a programming language. It includes functions. Shell code outside of functions is effectively code in an implicit main() function. An entire function should be fully seen on one page (~70-90 lines) and be readily comprehensible. If you have any doubts, then it is too complicated. Make it easier to understand by separating it into subroutines.
- The total length of a line (including comment) must not exceed 80 characters. Take advantage of bash's
+=operator for constants or linefeed escapes\.- Lines can be split without the need for a linefeed escape after
|,||,&and&&operators at the end of the line.
- Lines can be split without the need for a linefeed escape after
- The indentation must use 8-column tabs and not spaces. For line continuation, an additional tab should be used to indent the continued line, or align after
[or(for continued logic operations. - Comments are just as important in a shell script as in C code.
- Use
$(...)instead of`...`for subshell commands:$(...)is easier to see the start and end of the subshell command$(...)avoids confusion between'...'and`...`with a small font$(...)can be nested
- Use the subshell syntax only when necessary:
- When you need to capture the output of a separate program
- Using the construct with functions leads to stray output and/or convoluted code struggling to avoid output pollution
- It is more computationally efficient to not fork() the Bash process. Bash is slow enough already.
- Use "here string" like
function <<<$varinstead ofecho $var | functionto avoid forking a subshell and pipe - Use file arguments like
awk '...' $fileor input redirection likefunction << $fileinstead of a useless use ofcat - Use built-in Bash Parameter Expansion for variable/string manipulation rather than forking
sed/tr:- Use
${VAR#prefix}or${VAR%suffix}to removeprefixorsuffixrespectively - Use
${VAR/pattern/string}to replacepatternwith string instead ofecho $VAR | sed
- Use
- Avoid use of "
grep foo | awk '{ print $2 }'" since "awk '/foo/ { print $2 }'works just as well and avoids a separate fork + pipe - If a variable is intended to be used as a boolean, then it must be assigned as follows:
local mybool=false # or true
if $mybool; then
do_stuff
fi
- for loops it is possible to avoid a subshell for
$(seq 10)using the built-in iterator for fixed-length loops:- Unfortunately,
{1..$var}does not work, so use the(( ... ))arithmetic operator
- Unfortunately,
for ((i=0; i < $var; i++)); do
something_with $i
done
- Use
export FOOBAR=valinstead ofFOOBAR=val; export FOOBARfor clarity and simplicity - Use
[[ expr ]]for bash instead of[ expr ]- The
[[test understands regular expression matching with the=~operator - The easiest way to use it is by putting the expression in a variable and expanding it after the operator without quotes.
- The
- Use
(( expr ))instead of[ expr ]orlet exprwhen evaluating numerical expressions- This can include mathematical operators like
$((...)) - Will properly compare numeric values, unlike
...that is comparing strings
- This can include mathematical operators like
# wrong: this is surprisingly "true" because '5' > '3' and the rest of the string is ignored
[[ 5 > 33 ]] && echo "y" || echo "n"
y
# right: this is what you expect for the '>' operator
(( 5 > 33 )) && echo "y" || echo "n"
n
- This uses normal
<=,>=,==comparisons instead of-lt,-eq,-gt - Can use
for (( i=0; i <= END; i++ )); doand other numerical expressions instead of an external subshell forseq
- This uses normal
- Use
$((...))for arithmetic expressions instead ofexpr ...- No need for
$when referencing variable names inside$((...)) $((...))can handle hex values and common math operators
- No need for
- Error checks should prefer the form
[[ check ]] || actionto avoid leaving a dangling "false" on the return stack- Otherwise,
[[ check ]] && actionwill leave a dangling "false" on the stack ifcheckfails and an immediately following return/end of function will return an error
- Otherwise,
Test Framework
Variables
- Names of variables local to current test function which are not exported to the environment should be declared with "
local" and use lowercase letters - Names of global variables or variables that are exported to the environment should be UPPERCASE letters
- Use
$TMP/for temporary non-Lustre files instead of/tmp/ - Use
$SECONDSto get the current time when measuring test durations instead of$(date +%s)fork+exec:
local start=$SECONDS
do something
local elapsed=$((SECONDS - start))
or
local end=$((SECONDS + delay))
while ((SECONDS < end)); do
something
done
Functions
- Each function must have a section describing what it does and explain the list of parameters
# One line description of this function's purpose
#
# More detailed description of what the function is doing if necessary
#
# usage: function_name [--option argument] {required_argument} ...
# option: meaning of "option" and its argument
# required_argument: meaning of "required_argument"
#
# expected output and/or return value(s)
- Function arguments should be given local variable names for clarity, rather than being used as
$1 $2 $3in the function
local facet=$1
local file="$2"
local size=$3
- Use
sleep 0.1instead ofusleep 100000, sinceusleepis RHEL-specific
Tests and Libraries
- To avoid clustering a single
test-framework.shfile, there should be a<test-lib>.shfile for each test that contains specific functions and variables for that test. - Any functions, variables that global to all tests should be put in
test-framework.sh - A test file only need to source
test-framework.shand necessary<test-lib>.shfile
Subtests
- subtest names/numbers are not strictly meaningful in themselves, but there are some conventions that should be followed
- prefer subtest numbers below 1000 by convention
- some wrapper scripts (e.g. sanity-compr.sh) run other scripts like sanity.sh and sanityn.sh as well as their own subtests
- prefer subtest numbers below 1000 by convention
- it is convenient to cluster related tests together when adding new subtests
- easier to find related subtests to determine if there is coverage for some functionality
- adding subtests in different parts of the script (instead of at the end) avoids needless context patch conflicts
- adding subtests with disjoint numbers avoids semantic test numbering conflicts
- when adding subtests at the end of the file it is convenient to leave occasional gaps in the numbering (e.g. 100, 150, 200 instead of 33, 34, 35,)
- this leaves room for other subtests to be added with fewer context/number conflicts
- it is convenient to cluster related tests together when adding new subtests
- do not mix pure numbered subtests with suffixes for the same number (e.g. 56 and 56a)
- test-framework.sh allows running groups of subtests together (e.g. all 56X subtests), but if there is also plain-numbered subtest (e.g. 56) then it cannot be run in isolation
- if a numbered subtest is growing a new variant, rename the original subtest to the
asuffix (e.g. 56->56a) then add the new variant with thebsuffix - if the plain a-z suffixes are occupied or you want to make a variant of an already-suffixed subtest, use
aa,ab(oreaor similar) and locate it aftera(ore) - continue a-z suffixes with A-Z, but do not add aa, ab, etc. after the z or Z suffix to avoid clashes of subtest names
- subtests should not reference their test number explicitly, since the subtest name may change in the future (for various reasons).
- the
$testnumvariable holds the test number (e.g.11or27M) - the
$TESTNAMEvariable holds the test name (e.g.test_27)
- the
- test files should be named
$tfilefor the filename, or base name like$tfile.1or$tfile.sourceto simplify debugging - test directories should be named
$tdir, and should be used insetad of$DIRif more than a handful files are created for the subtest - small/few test files/dirs do not need to be explicitly deleted at the end of the test, that is done by test-framework.sh at the start/end of each test script
- large (over 1MB)/many (over 50) test files/dirs in a subtest should be cleaned up explicitly with a
stack_trapso that they are always cleaned up even if the test exits with an error, and do not fill the test filesystem
stack_trap "rm -f $DIR/$tfile.big"
fallocate -l 100M $DIR/$tfile.big || error "$tfile.big create failed"
stack_trap "unlinkmany $DIR/$tdir/$tfile- 1000"
createmany -o $DIR/$tdif/$tfile- 1000 || error "$tfile creation failed"
- creating large test files is by far the fastest with "fallocate" *when supported* (ldiskfs only), as determined by
check_set_fallocate- fallocate'd files will only produce zeroes when read, so it is not useful where the data contents are used since it is identical to a sparse file except by allocated blocks
- use
test_mkdirto add some variety to directory creation (random local, striped, remote) if directory location is not critical to the test - ensure that directory location and MDS facet are aligned. Since 2.14.54 directories may be created on any MDT, so "
do_facet mds1 ..." may be on the wrong MDS. - Use "
mkdir_on_mdt0 $DIR/$tdir" if necessary to create directories on MDT0000 for use withmds1, or preferentially "$LFS getdirstripe -m $DIR/$tdir" to determine MDT index, and "mds$((idx+1))" for facet name. - the
errormessages in a subtest should be unique so that it is easy to determine which check failed
lfs migrate -c3 $tfile || error "'lfs migrate -c3' failed"
lfs migrate -c1 $tfile || error "second 'lfs migrate -c1' failed"
- use
skipto skip subtests that should not run because of permanent functional deficiency (e.g. non-existent functionality in backing filesystem, older version of client/server, wrong configuration)
(( MDS1_VERSION_CODE >= $(version_code 2.15.53) )) ||
skip "need MDS >= 2.15.53 to check foobar works"
[[ $mds1_FSTYPE == "ldiskfs" ]] || skip "MDT0000 is not ldiskfs"
- use
skip_envfor minor environmental deficiency in developer test environment (e.g. missing binary) that _should_ exist in autotest:
kinit || skip_env "Kerberos not installed"
- when subtests are run in a loop with
ONLY_REPEAT=NorONLY_MINUTES=Mthen$ONLY_REPEAT_ITERholds the iteration number (1-based). In some cases this can be useful for adapting the test case to handle running many iterations in a row (e.g. addingwait_delete_completed) that isn't needed when run in isolation.