# Out of memory kill-error

**URL:** https://cms-helpdesk.ncas.ac.uk/t/out-of-memory-kill-error/1782
**Category:** Unified Model
**Created:** [21 August 2025 08:52 UTC](https://cms-helpdesk.ncas.ac.uk/t/out-of-memory-kill-error/1782 "2025-08-21T08:52:29Z")
**Posts on this page:** 4
**Page:** 1

<div class="post-metadata">

### Author: ![miclen](https://avatars.discourse-cdn.com/v4/letter/m/3da27b/32.png) [@miclen](https://cms-helpdesk.ncas.ac.uk/u/miclen)
#### Post date: [21 August 2025 08:52 UTC](https://cms-helpdesk.ncas.ac.uk/t/out-of-memory-kill-error/1782/1 "2025-08-21T08:52:29Z")

</div>

Hello, I’m trying to run a rather large domain (2000x2000 points) and when the run reaches the forecast stage, I’m encountering this error -

slurmstepd: error: Detected 1 oom-kill event(s) in StepId=10650430.0. Some of your processes may have been killed by the cgroup out-of-memory handler.  
srun: error: nid001161: task 0: Out Of Memory  
srun: launch/slurm: \_step\_signal: Terminating StepId=10650430.0  
slurmstepd: error: \*\*\* STEP 10650430.0 ON nid001161 CANCELLED AT 2025-08-21T09:34:37 \*\*\*  
slurmstepd: error: Detected 1 oom-kill event(s) in StepId=10650430.0. Some of your processes may have been killed by the cgroup out-of-memory handler.  
[FAIL] um-atmos \<\<‘ **STDIN** ’  
[FAIL]  
[FAIL] ‘ **STDIN** ’ # return-code=1  
2025-08-21T08:34:40Z CRITICAL - failed/EXIT

Is there a way to increase the memory available for the run? I’ve looked at some similar tickets posted, but I’m not sure if/where I could change the memory settings.

Best,

Michelle Maclennan

---

<div class="post-metadata">

### Author: ![RosalynHatcher](https://dub1.discourse-cdn.com/flex013/user_avatar/cms-helpdesk.ncas.ac.uk/rosalynhatcher/32/11_2.png) [@RosalynHatcher](https://cms-helpdesk.ncas.ac.uk/u/RosalynHatcher)
#### Post date: [21 August 2025 17:15 UTC](https://cms-helpdesk.ncas.ac.uk/t/out-of-memory-kill-error/1782/2 "2025-08-21T17:15:52Z")

</div>

Hi Michelle,

The standard compute nodes on ARCHER2 have 256Gb of memory and jobs have exclusive use of a compute node so you can’t request more memory.

You have 2 options:

1. There are some high memory nodes with 512Gb so you could try running on them. You access these by specifying the highmem partition & qos - see  
[Running jobs - ARCHER2 User Documentation](https://docs.archer2.ac.uk/user-guide/scheduler/#partitions)

2. Run on the standard nodes but underpopulated. How to do this is also detailed in the above web page.

Regards,  
Ros.

---

<div class="post-metadata">

### Author: ![miclen](https://avatars.discourse-cdn.com/v4/letter/m/3da27b/32.png) [@miclen](https://cms-helpdesk.ncas.ac.uk/u/miclen)
#### Post date: [26 August 2025 09:32 UTC](https://cms-helpdesk.ncas.ac.uk/t/out-of-memory-kill-error/1782/3 "2025-08-26T09:32:33Z")

</div>

Hi Ros,

Thank you for the advice! Is this done by changing ARCHER\_QUEUE=‘standard’ in the ./roses/suitname/rose-suite.conf file, or is there another file for the slurm submission script?

Best,

Michelle

---

<div class="post-metadata">

### Author: ![system](https://europe1.discourse-cdn.com/flex013/uploads/cms_support/original/1X/1fd2411499ffcbc299fe756cd5cdf26e44956558.png) [@system](https://cms-helpdesk.ncas.ac.uk/u/system)
#### Post date: [25 September 2025 09:33 UTC](https://cms-helpdesk.ncas.ac.uk/t/out-of-memory-kill-error/1782/4 "2025-09-25T09:33:00Z")

</div>

This topic was automatically closed 30 days after the last reply. New replies are no longer allowed.
