#645: Error concerning missing -V option when submitting job on LoneStar
------------------------+---------------------------------------------------
Reporter: hinder | Owner: eschnett
Type: defect | Status: new
Priority: major | Milestone: ET_2011_10
Component: SimFactory | Version:
Keywords: |
------------------------+---------------------------------------------------
I used the following command to submit an ET testsuite job on LoneStar.
The intention is to run on 1 process with 6 threads.
sim --remote lonestar create-submit maxwell_1proc_2 --testsuite --procs
6 --num-threads 6 --walltime 4:00:00 --ppn-used 6
and got a weird error. SimFactory didn't report any error or return a
nonzero exit code, even though it was unable to determine a job ID. I
repeated the submission, and the second time it worked, so the fault
appears to be intermittent. There was no difference in the submit script
in each case, apart from the job name.
The log file is attached, but the final error message from the log file
is:
{{{
[LOG:2011-10-22 11:29:05] self.submit(submitScript)::Executing submission
command: qsub
/scratch/00915/hinder/simulations/maxwell_1proc/output-0000/SIMFACTORY/SubmitScript
[LOG:2011-10-22 11:29:05] self.makeActive()::Simulation maxwell_1proc with
restart-id 0 has been made active
[LOG:2011-10-22 11:29:06] job_id = self.extractJobId(output)::received raw
output: Unable to run job: JSV rejected job.
[LOG:2011-10-22 11:29:06] job_id = self.extractJobId(output)::Exiting.
[LOG:2011-10-22 11:29:06] job_id =
self.extractJobId(output)::-----------------------------------------------------------------
[LOG:2011-10-22 11:29:06] job_id = self.extractJobId(output)::-- Welcome
to the Lonestar4 Westmere/QDR IB Linux Cluster --
[LOG:2011-10-22 11:29:06] job_id =
self.extractJobId(output)::-----------------------------------------------------------------
[LOG:2011-10-22 11:29:06] job_id = self.extractJobId(output)::--> Checking
that you specified -V...
[LOG:2011-10-22 11:29:06] job_id =
self.extractJobId(output)::--------------------------> Rejecting job
<--------------------------
[LOG:2011-10-22 11:29:06] job_id = self.extractJobId(output)::-V is now a
required option. Please specify it in your submit script.
[LOG:2011-10-22 11:29:06] job_id =
self.extractJobId(output)::---------------------------------------------------------------------
[LOG:2011-10-22 11:29:06] job_id = self.extractJobId(output)::
[LOG:2011-10-22 11:29:06] job_id = self.extractJobId(output)::using
submitRegex: Your job (\d+) \(.*?\) has been submitted
[LOG:2011-10-22 11:29:06] self.submit(submitScript)::After searching raw
output, it was determined that the job_id is: -1
[LOG:2011-10-22 11:29:06] self.submit(submitScript)::If this is -1, that
means the regex did NOT match anything. No job_id means no control.
}}}
Full log.txt file is attached. The job was not submitted.
The weird thing is that I do have -V in my submission script. The file
/scratch/00915/hinder/simulations/maxwell_1proc/output-0000/SIMFACTORY/SubmitScript
has
{{{
#! /bin/bash
#$ -A TG-MCA02N014
#$ -q normal
#$ -r n
#$ -l h_rt=4:00:00
#$ -pe 1way 12
#$
#$ -V
#$ -N maxwell_1proc-0
#$ -M ian.hinder(a)aei.mpg.de
#$ -m abe
#$ -o
/scratch/00915/hinder/simulations/maxwell_1proc/output-0000/maxwell_1proc.out
#$ -e
/scratch/00915/hinder/simulations/maxwell_1proc/output-0000/maxwell_1proc.err
cd /work/00915/hinder/Cactus/EinsteinToolkit
/work/00915/hinder/Cactus/EinsteinToolkit/simfactory/bin/sim run
maxwell_1proc --machine=lonestar --restart-id=0
}}}
Any ideas?
--
Ticket URL: <https://trac.einsteintoolkit.org/ticket/645>
Einstein Toolkit <http://einsteintoolkit.org>
The Einstein Toolkit
#690: CarpetIOASCII modification causes tests to fail
--------------------------------------+-------------------------------------
Reporter: hinder | Owner: eschnett
Type: defect | Status: new
Priority: major | Milestone:
Component: Carpet | Version:
Keywords: testsuites CarpetIOASCII |
--------------------------------------+-------------------------------------
The following changeset,
changeset: 3377:2f2e12bdf05b
tag: tip
user: Erik Schnetter <schnetter(a)cct.lsu.edu>
date: Tue Nov 15 11:59:53 2011 -0500
summary: CarpetIOASCII: Add new "compact" output format
causes the default CarpetIOASCII output format to change. Specifically,
{{{
# iteration 0
# refinement level 0 multigrid level 0 map 0 component 0 time
level 0
# column format: 1:it 2:tl 3:rl 4:c 5:ml 6:ix 7:iy 8:iz 9:time
10:x 11:y 12:z 13:data
0 0 0 0 0 0 18 3 0 -0.473684210526316 0.0967741935483871 0 0
}}}
becomes
{{{
# iteration 0 time 0
# time level 0
# refinement level 0 multigrid level 0 map 0 component 0
# column format: 1:it 2:tl 3:rl 4:c 5:ml 6:ix 7:iy 8:iz 9:time
10:x 11:y 12:z 13:data
0 0 0 0 0 0 18 3 0 -0.473684210526316
0.0967741935483871 0 0
}}}
This causes certain test cases (I think those which do not use
IO::out_fileinfo = "none") to fail. The problem is not in the
modification of the comment lines, whose content is ignored by the test
mechanism, but in the addition of a blank line or a comment line (the
"time level 0" in the above example). The test mechanism compares files
line-by-line, and so becomes out of step if additional comment or blank
lines are introduced.
The attached patch modifies the test mechanism to completely ignore the
presence of any blank or comment lines. With this patch, the tests
bhns_eval, bhns_interp, checkpointML and recoverML now pass again.
Is this solution acceptable? Should the test mechanism be modified in a
different way? Or should CarpetIOASCII be modified to produce the old
format?
--
Ticket URL: <https://trac.einsteintoolkit.org/ticket/690>
Einstein Toolkit <http://einsteintoolkit.org>
The Einstein Toolkit
#689: Tests which use ML_ADMConstraints need to have results regenerated
-----------------------------------+----------------------------------------
Reporter: hinder | Owner:
Type: defect | Status: new
Priority: major | Milestone:
Component: EinsteinToolkit thorn | Version:
Keywords: ML_ADMConstraints |
-----------------------------------+----------------------------------------
The correction implemented in #671 causes tests which have the incorrect
values in their output files to fail. These need to be regenerated.
Waiting for other test failures to be understood before doing this,
however.
--
Ticket URL: <https://trac.einsteintoolkit.org/ticket/689>
Einstein Toolkit <http://einsteintoolkit.org>
The Einstein Toolkit
#635: Cactus should print the number of processes used when running tests
-------------------------+--------------------------------------------------
Reporter: hinder | Owner:
Type: enhancement | Status: new
Priority: minor | Milestone:
Component: Cactus | Version:
Keywords: |
-------------------------+--------------------------------------------------
When running a test suite, it is usually important to know how many MPI
processes are being used, as this determines which tests are run. This
information is not currently present in the summary.log file, and is not
output to standard output either. This makes it hard to check that
simfactory is actually running a test on the expected number of processes.
The attached patch outputs the number of processes to these two places.
OK to apply before the release?
--
Ticket URL: <https://trac.einsteintoolkit.org/ticket/635>
Einstein Toolkit <http://einsteintoolkit.org>
The Einstein Toolkit
#668: Cactus does not strictly enforce STEERABLE values
--------------------+-------------------------------------------------------
Reporter: knarf | Owner:
Type: defect | Status: new
Priority: minor | Milestone: Cactus_4.1.0
Component: Cactus | Version: development version
Keywords: |
--------------------+-------------------------------------------------------
Cactus currently checks for valid values of STEERABLE (parameters) by
looking for a sub-string, not an exact match. This way, e.g., RECOVERY is
treated like RECOVER. The attached patch fixes this. Apart from two
private changes this doesn't affect public thorns (I compiled, but that
list is long), and in case it does a clear error message is printed.
--
Ticket URL: <https://trac.einsteintoolkit.org/ticket/668>
Einstein Toolkit <http://einsteintoolkit.org>
The Einstein Toolkit
#675: allow zero timelevels in STORAGE when timelevels are specified via
parameter
-------------------------+--------------------------------------------------
Reporter: rhaas | Owner:
Type: enhancement | Status: new
Priority: minor | Milestone:
Component: Cactus | Version:
Keywords: |
-------------------------+--------------------------------------------------
this allows to simply set timelevels=0 to turn off storage without having
to put an "if(do_something || timelevels > 0)" into schedule.ccl files. It
is also the only way to turn off storage inside of a GROUP of SCHEDULE
statement based on a condition (other than scheduling the item twice, once
with STORAGE, once without).
More complicated conditions require either (I believe):
a. the (ab)use of accumulors the way that GRHydro::GRHydro_hydro_excision
works (only all "and" or all "or" supported I think)
a. using undocumented behaviour and putting C code into schedule.ccl
{{{
const int compound_condition = condition_a || condition_b ? 3 : 2
SCHEDULE
{
STORAGE foo[compound_condition]
...
}}}
c. allow more than just identifiers within the '[]" brackets of STORAGE
(most likely one has to allow almost everything "{{{[^]]+}}}") then use
C's "?" operator to build up the complicated expression
Having 0 timelevels works trivially since eventually these STORAGE
statements end up in CCTK_GroupStorageIncrease which allows 0 timelevels.
--
Ticket URL: <https://trac.einsteintoolkit.org/ticket/675>
Einstein Toolkit <http://einsteintoolkit.org>
The Einstein Toolkit
#691: CarpetIOASCII tests "compact" and "coords" have nans in their output
--------------------------------------+-------------------------------------
Reporter: hinder | Owner: eschnett
Type: defect | Status: new
Priority: major | Milestone:
Component: Carpet | Version:
Keywords: CarpetIOASCII testsuites |
--------------------------------------+-------------------------------------
The spherical surface output files in these tests contain nans. Was this
deliberate to test the output of nans? When the tests are run on datura,
the output files contain "-nan" instead of "nan", causing the tests to
fail.
--
Ticket URL: <https://trac.einsteintoolkit.org/ticket/691>
Einstein Toolkit <http://einsteintoolkit.org>
The Einstein Toolkit
#683: LSUThorns/Vectors: Remove pos, add sin/cos/tan functions
-----------------------------------+----------------------------------------
Reporter: eschnett | Owner:
Type: enhancement | Status: new
Priority: major | Milestone:
Component: EinsteinToolkit thorn | Version:
Keywords: |
-----------------------------------+----------------------------------------
Remove kpos, because it is not used (it is a no-op, i.e. the
arithmetic + operator).
Add sin, cos, and tan.
--
Ticket URL: <https://trac.einsteintoolkit.org/ticket/683>
Einstein Toolkit <http://einsteintoolkit.org>
The Einstein Toolkit
#223: Make declaration of CCTK_ARGUMENTS safer
-------------------------+--------------------------------------------------
Reporter: eschnett | Owner:
Type: enhancement | Status: new
Priority: minor | Milestone:
Component: Cactus | Version:
Keywords: |
-------------------------+--------------------------------------------------
I suggest to change the declaration of CCTK_ARGUMENTS in C from
cGH * cctkGH
to
cGH const * CCTK_RESTRICT const cctkGH
which should lead to safer code and may even enable some optimisations.
--
Ticket URL: <https://trac.einsteintoolkit.org/ticket/223>
Einstein Toolkit <http://einsteintoolkit.org>
The Einstein Toolkit