Known issues

Unresolved known issues

Known issue with an Unresolved Resolution state is an active problem under investigation; a temporary workaround may be available.

Resolved known issues

A known issue with a Resolved (workaround) Resolution state is an ongoing problem; a permanent workaround is available which may include using different software or hardware.

A known issue with Resolved Resolution state has been corrected.

Known Issues

Title Category Resolutionsort descending Description Posted Updated
MVAPICH broken on Ruby Ruby Resolved

Update Monday February 16th -- Ruby MVAPICH2 build fixed.

Ruby's MVAPICH2 build has been fixed.  Please email oschelp@osc.edu with any issues.

... Read more
11 years 6 months ago 11 years 5 months ago
Handling full-node MPI warnings with MVAPICH 3.0 Ascend, Cardinal, Pitzer, Software Resolved
(workaround)

When running a full-node MPI job with MVAPICH 3.0 , you may encounter the following warning message:

[][mvp_generate_implicit_cpu_mapping] WARNING: You appear to be...
Read more
1 year 9 months ago 6 months 6 days ago
Application Errors client portal Resolved

When beginning a major or discovery-level application for resources at OSC, you are asked for a required justification on the Additional Documents page. However, there is no mechanism for you to...

Read more
7 years 3 months ago 7 years 1 week ago
Segmentation fault from openmpi/1.10-hpcx and 2.0-hpcx on Owens Owens, Software Resolved

We have found that recent MPI jobs using openmpi/1.10-hpcx and openmpi/2.0-hpcx on Owens may complete or hang until the job is killed, but receive segmentation fault. Some applications might be ...

Read more
7 years 2 weeks ago 6 years 12 months ago
MyOSC budget balance may not be correct client portal Resolved

Resolved

Version 3.0.1 was deployed which patches this issue. View the changelog for details.

Original post

...

Read more
4 years 5 months ago 3 years 11 months ago
Rolling reboot of all clusters, starting from Wednesday morning, April 19, 2017 Batch, Maintenance, Owens, Ruby Resolved

1:40PM 4/27/2017 Update: Rolling reboots are completed. 

3:10PM 4/18/2017 Update: Rolling reboots on Owens have started to address GPFS errors occured...

Read more
9 years 3 months ago 9 years 3 months ago
Rolling Reboot for Security Fix Resolved

A rolling reboot is in progress to address CVE-2026-23111 (nf_tables logic bug) for all clusters, including Ascend, Cardinal, and Pitzer. Login nodes will be rebooted first and access...

Read more
1 month 4 weeks ago 1 month 2 weeks ago
Email Issues client portal Resolved

OSU is having ongoing periodic problems with Microsoft (their mail hosting provider) severely delaying outbound email. There is no solution being offered and no timeline for getting it resolved....

6 years 10 months ago 6 years 6 months ago
Lustre is still offline. HPC systems back up Maintenance Resolved

Day One of the scheduled downtime has been completed, and HPC operations have resumed. As planned, Lustre work will extend into Day Two. Jobs using /fs/lustre or $PFSDIR cannot run until this work...

Read more
12 years 1 month ago 12 years 1 month ago
Slurm database repair on 01/25/2024 Outage Resolved

We have scheduled a Slurm database repair, which is planned to start at 8:30 am US/Eastern on Thursday, January 25, 2024. During the repair, Slurm database will be offline; running jobs and...

Read more
2 years 6 months ago 2 years 6 months ago

Pages