Hi,
I'm getting out of memory (oom) errors on a system with the following specs:
- JIRA Server
- Jira Core 7.7.1
- ~10 users
- ~20 projects
- ~1.5 million issues
- AWS mx4.2xlarge 8cpu, 32GB Ram
- PostgreSQL colocated
- Xmax=16384Mb
We have a large import job--1000s of new issues--and subsequent REST API-driven updates with built-in 15 second delays. We started hitting ooms after 75 minutes, and have bumped up Xmax from 2GB to 3, then 4 then 8, now 16. We were failing with ooms after 75 minutes at first. Now, it takes a 3-4 hours before CPU utilization jumps to ~80% and GC errors multiply- and eventually ooms.
I've looked at the sizing recommendations, but all are based on scaling multiple dimensions simultaneously: issues, users, projects, etc. We only have issues at scale.
I'm digging deeper into plugins, indexing, searching performance as suggested by other posts, but i'm hoping there is an obvious thing to try, like "Oh yeah, issue count is the dominant scaling factor because of xyz so double your instance size to 16/64". Or "tune your gc according to recommendations in this document" or "it's definitely a search problem, try to optimize that".
Please note that we have been steadily scaling with this process for over two years, and have only experienced this problem two or three times, however this week it's been in the headlines consistently.
Thanks!