Site is running well
- Full of Atlas jobs (MCORE, SCORE, Analy and Opport)
- Some problems with high mem MCORE last week
- Using 3-3.5GB per core RSS
- Excessive swapping on lower memory nodes
- Inefficient
- Caused GPFS problems on low mem Illinois nodes
- These jobs are now gone
Adjust space token
- Armen pointed out that LOCALGROUPDISK was nearly full
- Used the last of the retired PRODDISK + old OSG space token (~40TB)
- May need to move more from USERDISK and GROUPDISK
MWT2 Accounting verified
- WLCG numbers (from OSG Gratia) are very close to Panda (larger by ~2%)
- Panda numbers come from Atlas Dashboard
- Difference from local Tier3 AtlasConnect jobs running on Tier2 cores.
Networking
- Illinois ⇔ Indiana high latency
- Taking inefficient path via I2
- Illinois networking is aware and will work with IndianaGigaPOP
- Need to advertise direct route between two sites via OmniPOP
New hardware status
- UChicago
- 18 Ceph Servers
- Rolling online of new servers / upgrade of existing servers.
- Indiana
- 24 R630 (E5-250, 128GB) have been delivered
- Waiting on PDU to provide power
- Illinois - Done