Known issues

Unresolved known issues

Known issue with an Unresolved Resolution state is an active problem under investigation; a temporary workaround may be available.

Resolved known issues

A known issue with a Resolved (workaround) Resolution state is an ongoing problem; a permanent workaround is available which may include using different software or hardware.

A known issue with Resolved Resolution state has been corrected.

Known Issues

Title Category Resolutionsort descending Description Posted Updated
Ls-dyna license outage since Oct 15, 2019 Licensing Resolved

Updated on 1:47 PM Oct 16, 2019

The ls-dyna license server is operational again from 1:40 pm on Oct 16, 2019

Original Post:

The ls-dyna...

Read more
6 years 11 months ago 6 years 11 months ago
Rolling reboot of Owens cluster, starting from 9AM September 11, 2017 Batch, Owens Resolved

Updates on 12:20PM September 25, 2017: 

The rolling reboot of Owens is completed. 

... Read more
9 years 1 week ago 8 years 11 months ago
Fail to connect with VS Code 1.86 Resolved

VS Code 1.86 (aka the ‘January 2024’ update) requires ≥glibc 2.28 which are not supported on Pitzer and Owens clusters. Please downgrade to VS Code 1.85. 

See this link for more...

Read more
2 years 7 months ago 2 years 4 months ago
Statewide Intel compiler license checkout failures Licensing Resolved

This morning (9/10/14) we updated our Intel compiler licenses. We are seeing some unexpected license checkout failures in the logs (please click through to see details):

10:44:...
Read more
12 years 5 days ago 11 years 11 months ago
Rolling Reboot of all three HPC clusters beginning July 7 Resolved

A rolling reboot is scheduled to remediate a ptrace vulnerability and safely restore debugger functionality for all clusters, including Ascend, Cardinal, and Pitzer...

Read more
2 months 2 weeks ago 1 month 3 weeks ago
GPFS errors on compute nodes filesystem Resolved

We've seen an increase in transient problems that result in compute nodes losing access to the GPFS file systems for ~5 minutes.

Any jobs running on these nodes accessing files on GPFS may...

Read more
5 years 9 months ago 4 years 9 months ago
OnDemand service is not available OnDemand Resolved

Update:

OnDemand service is available again.

Original post:

OnDemand service is not available and won't let users log in. We are working to fix it as soon as we can....

Read more
7 years 8 months ago 7 years 8 months ago
PyTorch hangs on dual-gpu node on Ascend Ascend, GPU Resolved
(workaround)

PyTorch can hang on Ascend on dual-GPU nodes

Through internal testing, we have confirmed that the hang issue only occurs on Ascend dual-GPU (nextgen) nodes. We’re still unsure why...

Read more
1 year 4 months ago 1 year 4 months ago
Submit filter bug after downtime Batch Resolved

A change was made to a part of our batch software during the downtime that should have only affected users who are a part of multiple projects. We have found that there is a bug in the changes...

Read more
10 years 7 months ago 10 years 7 months ago
Oakley login nodes and ruby02 will not be accessible between 9:00-9:30am on 10/18/2016 login Resolved

We upgraded to RHEL 6.8 for both Oakley and Ruby clusters during the October 12th's downtime. Unfortunately, we are noticing some NFS problem that has been causing rsh, or ssh sessions to hang on...

Read more
9 years 11 months ago 9 years 11 months ago

Pages