Known issues

Unresolved known issues

Known issue with an Unresolved Resolution state is an active problem under investigation; a temporary workaround may be available.

Resolved known issues

A known issue with a Resolved (workaround) Resolution state is an ongoing problem; a permanent workaround is available which may include using different software or hardware.

A known issue with Resolved Resolution state has been corrected.

Known Issues

Title Category Resolutionsort ascending Description Posted Updated
Backup failures on ess filesystem Backups, filesystem Resolved

The backups on the /fs/ess filesystem are having issues running. There has not been a successful backup of this filesystem since Sunday, 08 August 2021.

OSC is working with the vendor to...

Read more
5 years 1 week ago 5 years 1 week ago
Data on /fs/scratch is not accessible filesystem Resolved

Updated on 10:30 AM July 3rd, 2019:

Data on /fs/scratch is accessible now. We are working with the vendor to find the root cause and apologize for any inconvenience.  ...

Read more
7 years 1 month ago 7 years 1 month ago
Inconsistent performance degradation of ESS filesystem filesystem Resolved

...

Read more
4 years 8 months ago 4 years 7 months ago
Brief interruption of batch services on 4/17 Batch Resolved

On April 17th 2013, at roughly 2PM, we will be rebooting the batch server on the Oakley cluster. Running jobs will not be affected, but there will be a brief disruption in scheduling, as well as...

Read more
13 years 4 months ago 13 years 4 months ago
cuMemHostRegister Fails with CUDA_ERROR_INVALID_VALUE on RHEL 9.6 Ascend, Cardinal, GPU, system software Resolved

After upgrading the operating system to RHEL 9.6 during the scheduled downtime on May 12, 2026,  applications utilizing UCX (...

Read more
2 months 3 weeks ago 1 month 3 weeks ago
Peer Review Submissions client portal Resolved

When submitting a peer review, an error message appears:

Error during rendering of region "...

Read more
6 years 10 months ago 6 years 10 months ago
Rolling reboot of Owens cluster, starting from 9AM June 28, 2017 Owens Resolved

Update posted on July 7, 2017 at 2:00PM:

Rolling reboot of login and compute nodes of Owens cluster is completed. 

... Read more
9 years 1 month ago 9 years 1 month ago
Running jobs requeued on all clusters Owens, Pitzer Resolved

The Slurm upgrades during rolling reboots of Ascend, Owens and Pitzer we performed today (Oct 25 2023) cause all running jobs on the systems requeued around 8:45am. You will not be billed for the...

Read more
2 years 9 months ago 2 years 9 months ago
Lustre jobs suspended filesystem Resolved

The Lustre filesystem ($PFSDIR and /fs/lustre) has crashed several times Friday evening (8/15). We have degraded this service temporarily, while we work to isolate the actions that are triggering...

Read more
12 years 1 week ago 11 years 12 months ago
OSC internal network problems 25 Sept. 2020 Network, Owens, Pitzer, Ruby, Web Services Resolved

OSC is currently experiencing problems with its internal network.  Interactive sessions may be slow or unresponsive, but running jobs should not be affected.

5 years 10 months ago 5 years 10 months ago

Pages