Nested domain

Akash6139

New member
Sir
I was running the model at 9,3,1 km resolution but its not running it stopped after some time withing seconds and this is my namelist.input


cat namelist.input

&time_control
start_year = 2025, 2025, 2025,
start_month = 06, 06, 06,
start_day = 29, 29, 29,
start_hour = 00, 00, 00,
start_minute = 00, 00, 00,
start_second = 00, 00, 00,
end_year = 2025, 2025, 2025,
end_month = 07, 07, 07,
end_day = 02, 02, 02,
end_hour = 06, 06, 06,
end_minute = 00, 00, 00,
end_second = 00, 00, 00,
interval_seconds = 10800
input_from_file = .true.,.true.,.true.,
history_interval = 60, 60, 60,
frames_per_outfile = 1, 1, 1,
restart = .false.,
restart_interval = 7200000,
io_form_history = 2
io_form_restart = 2
io_form_input = 2
io_form_boundary = 2
debug_level = 0
write_input = .false.,
inputout_interval = 180, 360,
inputout_begin_h = 3, 6,
inputout_end_h = 3, 6,
input_outname = "wrfvar_input_d03_2025-07-02_06:00:00",
/


&domains
time_step = 36,
time_step_fract_num = 0,
time_step_fract_den = 1,
max_dom = 3,
e_we = 125, 202, 391,
e_sn = 125, 223, 448,
e_vert = 48, 48, 48,
dzstretch_s = 1.1
p_top_requested = 5000,
num_metgrid_levels = 34,
num_metgrid_soil_levels = 4,
dx = 9000,
dy = 9000,
grid_id = 1, 2, 3,
parent_id = 1, 1, 2,
i_parent_start = 1, 29, 36
j_parent_start = 1, 26, 37
parent_grid_ratio = 1, 3, 3,
parent_time_step_ratio = 1, 3, 3,
feedback = 1,
smooth_option = 1,
/

&physics
mp_physics = 6, 6, 6,
cu_physics = 1, 0, 0,
ra_lw_physics = 1, 1, 1,
ra_sw_physics = 1, 1, 1,
bl_pbl_physics = 1, 1, 1,
sf_sfclay_physics = 1, 1, 1,
sf_surface_physics = 2, 2, 2,
radt = 9, 3, 1,
bldt = 0, 0, 0,
cudt = 9, 0, 0,
icloud = 1,
num_land_cat = 21,
sf_urban_physics = 0, 0, 0,
fractional_seaice = 1,
/

&fdda
/

&dynamics
hybrid_opt = 2,
etac = 0.1,
w_damping = 1,
diff_opt = 2, 2, 2,
km_opt = 4, 4, 4,
diff_6th_opt = 0, 0, 0,
diff_6th_factor = 0.12, 0.12, 0.12,
base_temp = 290.
damp_opt = 3,
use_theta_m = 1,
zdamp = 5000., 5000., 5000.,
dampcoef = 0.2, 0.2, 0.2,
khdif = 0, 0, 0,
kvdif = 0, 0, 0,
non_hydrostatic = .true., .true., .true.,
moist_adv_opt = 1, 1, 1,
scalar_adv_opt = 1, 1, 1,
gwd_opt = 1, 0, 0,
/

&bdy_control
spec_bdy_width = 5,
specified = .true.
/



&grib2
/

&namelist_quilt
nio_tasks_per_group = 0,
nio_groups = 1,
/

For two domains, it is running, but when I introduce the third one, it stops. Please guide me further. Is there something wrong in the namelist section? or something else
 

HPC Job Terminating with SIGKILL (Signal 9)​

I am running a WRF job on an HPC cluster using MPI. The job starts normally, but during execution several MPI processes are terminated with Signal 9 (SIGKILL).

The job log shows errors such as:

===================================================================================<br>= BAD TERMINATION OF ONE OF YOUR APPLICATION PROCESSES<br>=<br>= RANK 78 PID 4173397 RUNNING AT cn1161<br>= KILLED BY SIGNAL: 9 (Killed)<br>===================================================================================<br><br>===================================================================================<br>= BAD TERMINATION OF ONE OF YOUR APPLICATION PROCESSES<br>=<br>= RANK 79 PID 4173398 RUNNING AT cn1161<br>= KILLED BY SIGNAL: 9 (Killed)<br>===================================================================================<br><br>===================================================================================<br>= BAD TERMINATION OF ONE OF YOUR APPLICATION PROCESSES<br>=<br>= RANK 80 PID 4173399 RUNNING AT cn1161<br>= KILLED BY SIGNAL: 9 (Killed)<br>===================================================================================<br><br>===================================================================================<br>= BAD TERMINATION OF ONE OF YOUR APPLICATION PROCESSES<br>=<br>= RANK 81 PID 4173400 RUNNING AT cn1161<br>= KILLED BY SIGNAL: 9 (Killed)<br>===================================================================================<br><br>===================================================================================<br>= BAD TERMINATION OF ONE OF YOUR APPLICATION PROCESSES<br>=<br>= RANK 82 PID 4173401 RUNNING AT cn1161<br>= KILLED BY SIGNAL: 9 (Killed)<br>===================================================================================<br><br>===================================================================================<br>= BAD TERMINATION OF ONE OF YOUR APPLICATION PROCESSES<br>=<br>= RANK 83 PID 4173402 RUNNING AT cn1161<br>= KILLED BY SIGNAL: 9 (Killed)<br>===================================================================================<br><br>===================================================================================<br>= BAD TERMINATION OF ONE OF YOUR APPLICATION PROCESSES<br>=<br>= RANK 95 PID 2422884 RUNNING AT cn1170<br>= KILLED BY SIGNAL: 9 (Killed)<br>===================================================================================<br>
The interesting point is that multiple MPI ranks are being killed on the same compute node (cn1161), while another rank is killed on cn1170.

There is no clear WRF-specific error message immediately before the termination.

​

Any guidance on diagnosing this issue would be appreciated.
 
Your namelist.input looks good. Where is your model domain? Can you post a figure to show the terrain over D01 and D03?

Please recompile WRF in debug mode, then run the case again. We need to know exactly when and where the model crashes first.
 
Back
Top