Status updates | sota staircase Incidents and maintenance reported on status page for sota staircase https://status.joshwel.co/ https://d1lppblt9t2x15.cloudfront.net/logos/da91c1bb3376bc30be582cd4981ef800.png Status updates | sota staircase https://status.joshwel.co/ en Maintenance: maintenance https://status.joshwel.co/maintenance/1079921 Fri, 02 Oct 2026 08:50:04 -0000 https://status.joshwel.co/incident/1079921#b80a1de04e7550eb79ffdba94586cef8d621c9af3ed45f19f01537ceaa416623 Maintenance Maintenance completed Maintenance: maintenance https://status.joshwel.co/maintenance/1079921 Fri, 02 Oct 2026 08:37:04 -0000 https://status.joshwel.co/incident/1079921#2a7851ab13c632b11535cbb369aaa82d540693ec4d58677fe9431131f8a8e4c3 Maintenance culling the server into 12c + powersave; may reboot Maintenance: this is the end https://status.joshwel.co/maintenance/1075847 Thu, 01 Oct 2026 18:02:07 -0000 https://status.joshwel.co/incident/1075847#3504f99e496640a60c1d1fb40dce823a6cd599d648c4a37c4a41ab6b9a1ad25d Maintenance and we up Maintenance: this is the end https://status.joshwel.co/maintenance/1075847 Thu, 01 Oct 2026 18:02:07 -0000 https://status.joshwel.co/incident/1075847#0c5c319c61f98bea38f9b0a31f6a4140132e886bab8aa92644e2bfb3aa10065f Maintenance Maintenance completed Maintenance: this is the end https://status.joshwel.co/maintenance/1075847 Mon, 28 Sep 2026 01:30:00 -0000 https://status.joshwel.co/incident/1075847#39b4e0b18c76b741323c7e6bc55fa4408ffec63dc2feab9ccce861711c20223b Maintenance ryan is renovating Maintenance: WE ARE GETTING A KVM! https://status.joshwel.co/maintenance/942469 Fri, 03 Jul 2026 09:53:00 -0000 https://status.joshwel.co/incident/942469#5344389a02e65144cb3ce254b734f831674c9b903acb11f86dad594ea75947b4 Maintenance so the server has to go DOWN to plug in mobo power control server went down, and we know https://status.joshwel.co/incident/933377 Thu, 25 Jun 2026 16:48:00 -0000 https://status.joshwel.co/incident/933377#6ba40ac850bd4918c7bf3ebdeb6196527177482b292b36a3395b1a422036ab28 Incident back online server went down, and we know https://status.joshwel.co/incident/933377 Thu, 25 Jun 2026 16:45:00 -0000 https://status.joshwel.co/incident/933377#77e545e7f10943b7fa9cc9e5fc1df668b90878688631ceb63aafea5112053648 Incident the fix has been applied, verified and the nixos system flake no longer hard-depends on md0 coming up; this should no longer happen with disastrous recovery setiup needed in the future server will be replaced in a cupboard and come back online soon server went down, and we know https://status.joshwel.co/incident/933377 Thu, 25 Jun 2026 16:13:00 -0000 https://status.joshwel.co/incident/933377#2aed698a8994935cb575f62a76e090ef5dbec417510c86dddf2a89953aebe363 Incident a reverted uuid change in the nixos hardware configuration has only shown itself to be erroneous due to the rebooting of the system during a power surge this uuid will be updated, and the system will be rebuilt from another live environment server went down, and we know https://status.joshwel.co/incident/933377 Wed, 24 Jun 2026 16:12:00 -0000 https://status.joshwel.co/incident/933377#ff2b3455cf257b5aecd32b225cd445d1481bcbcb8f75ec11b69e2a1cec0fec37 Incident nixos, the server's linux distro, is unable to find the bulk data raid array at boot however, it is visible and r/w-able in fedora thanks to livecd auto-assembling even if a disk has died, the 3x8TB hdds are set up in a RAID 10 mdadm array, meaning one drive failure will result in data survival server went down, and we know https://status.joshwel.co/incident/933377 Wed, 24 Jun 2026 11:20:00 -0000 https://status.joshwel.co/incident/933377#e9d4a3e6d3f43bbd2282c56b73998e3189db0d49134a11437d19a26be0a6afda Incident 🚨 server is down, and i know, and am fixing it affected services: all forgejo @ forge.joshwel.co affected users: n/a objectstore @ datashower.joshwel.co affected users: me and ryan immich @ i.scorchthinks.dev affected users: ryan-side + cmm album r/w jellyfin @ watch.scorchthinks.dev affected users: me and ryan hope to be back up by 1940 earliest, no upper bound on latest separating hot and bulk data https://status.joshwel.co/incident/931869 Tue, 23 Jun 2026 09:19:00 -0000 https://status.joshwel.co/incident/931869#3db9db07445478e68b7a2991397c63bf753a94094f8d3f743298db615fbd39a8 Incident services are back online separating hot and bulk data https://status.joshwel.co/incident/931869 Tue, 23 Jun 2026 05:20:00 -0000 https://status.joshwel.co/incident/931869#f25f27eea405e87162d60680e2fe531e770c4b9e51e4b873555da79e7e9fe47c Incident moving container dbs and configs into nvme to prevent service slowdowns; letting bulk data like repos, media, and objects take its time stretch goal is also docker root hopefully will drastically improve latency disks are saturated https://status.joshwel.co/incident/931779 Tue, 23 Jun 2026 05:19:00 -0000 https://status.joshwel.co/incident/931779#cdf6470ca89a1aa0c91f096e5da3e730aa63f782cd8951a359462c9f555ae51f Incident we good but not for long disks are saturated https://status.joshwel.co/incident/931779 Tue, 23 Jun 2026 03:19:00 -0000 https://status.joshwel.co/incident/931779#b217653314dda30fced85702441ce46b3314fcb5a485796370d9b97554e5a434 Incident seeding a little too much linux ISOs for poor SMR hdds trying out things; bringing media stack down, etc minecraft killed the ssd https://status.joshwel.co/incident/892466 Sun, 17 May 2026 16:35:00 -0000 https://status.joshwel.co/incident/892466#fa1f7e8bfc9b6e1b621fad00bd8a8143dcca2a96dcd88d9c9528b87db62df37f Incident disk saturation from emergency recovery of /mnt/ssd is done, we up minecraft killed the ssd https://status.joshwel.co/incident/892466 Sun, 17 May 2026 12:46:00 -0000 https://status.joshwel.co/incident/892466#03bc2ae739e6b94ee8dac7086788cbd3eae7687dc7fa58a8a8e3a0310575d83a Incident disks are saturated; server performance will be slow till est noon tomorrow, 1200 18/5/2026 UTC+8 (SGT) minecraft killed the ssd https://status.joshwel.co/incident/892466 Sun, 17 May 2026 10:12:00 -0000 https://status.joshwel.co/incident/892466#d6f6fdc9ee7fc78a8b2d31b13846abd7c0344483eff3121225c6b03461c8742b Incident should be up (except mc); monitoring stuff minecraft killed the ssd https://status.joshwel.co/incident/892466 Fri, 15 May 2026 18:42:00 -0000 https://status.joshwel.co/incident/892466#c1e11246cba93af8f0afb4d72c957af6e48a940e38d1546cdf07db0bfed01f69 Incident due to ryan and me having zero time, the restoration will be done on sunday minecraft killed the ssd https://status.joshwel.co/incident/892466 Sun, 10 May 2026 15:00:00 -0000 https://status.joshwel.co/incident/892466#2bee23e930330961e9d221cc883cbdb3f5ec05638c7d1e0b4202547398e60707 Incident server boots off an nvme, core services like the forge, immich, and object store are on a hdd raid array this only affects a personal minecraft instance but we're working on data recovery lets slime on mama https://status.joshwel.co/incident/797861 Sat, 10 Jan 2026 18:32:00 -0000 https://status.joshwel.co/incident/797861#f774a89b66e1ddebbdc3cb8b7c645503aba44f4bf21836cf4a3a1ffa1813c903 Incident we up lets slime on mama https://status.joshwel.co/incident/797861 Tue, 06 Jan 2026 11:45:00 -0000 https://status.joshwel.co/incident/797861#080084fea2f49a73e9a773f3c9ee21e7a26ecbc329443ba259849daa3a09adf8 Incident ocis is GONE buh bye there will be planned migration of data this is the reckoning thanks bye Maintenance: everything is DOWN https://status.joshwel.co/maintenance/799346 Tue, 06 Jan 2026 11:30:52 -0000 https://status.joshwel.co/incident/799346#ae90436e57680e6c8e066a4b29bc674d16f9c2fe8289a6e742d788e75b850c5a Maintenance we are going to migrate our main raid array from RAID5 to RAID10 (1+0) because ryan forgot the drives were SMR instead of CMR retort by ryan: "i did not know there was this distinction" retort-retort by mark: i forgot tbh Maintenance: system upgrade https://status.joshwel.co/maintenance/797859 Sat, 03 Jan 2026 11:30:07 -0000 https://status.joshwel.co/incident/797859#d055260df372b7603162ef02eb1f93a94d0e0fd80cd1a055c84d5ba6f2949916 Maintenance moving from NixOS 25.05 to NixOS 25.11, and will reboot afterwards lets slime on mama https://status.joshwel.co/incident/797861 Sat, 03 Jan 2026 09:40:00 -0000 https://status.joshwel.co/incident/797861#371530b8236bf86fb8a09c21373d5b088f99ec3f9bd09556f0ac0960e6fbe6e1 Incident disks are saturated the server is down, and now we know https://status.joshwel.co/incident/776282 Tue, 02 Dec 2025 15:59:00 -0000 https://status.joshwel.co/incident/776282#5c5303cc4d88b2eedf1a45f44be97f3225b6f9622947764762de37a2f362e764 Incident resolved the server is down, and now we know https://status.joshwel.co/incident/776282 Sat, 29 Nov 2025 08:05:00 -0000 https://status.joshwel.co/incident/776282#8adeeb039fb3d8b52659c6a2b428154df15ce39e9b4de965fae624bc7e89b968 Incident all bare metal services are currently unresponsive and time out; we are aware and will resolve it in the following hours. any affected users if at all may consider reaching out to mark at joshwel dot co for urgent data retrieval or likewise Maintenance: object store is being redeployed!! https://status.joshwel.co/maintenance/718938 Fri, 05 Sep 2025 11:27:25 -0000 https://status.joshwel.co/incident/718938#e0a0aa9c576371a1853e385c99119668b3536117c586dcd19f62dfb332d3c2be Maintenance was previously a monitored preprod/test instance but i like it!!! so i'm redeploying it with slightly different configs!!! Maintenance: IMMICH IS DOWN, MIGRATING DATA https://status.joshwel.co/maintenance/717219 Tue, 02 Sep 2025 16:00:22 -0000 https://status.joshwel.co/incident/717219#fece049681076622354babef39295cb0dc92d44044fc417a8146367a8517a112 Maintenance data is currently being migrated, disks are probably saturated Maintenance: EVERYTHING IS DOWN https://status.joshwel.co/maintenance/713946 Thu, 28 Aug 2025 03:00:00 -0000 https://status.joshwel.co/incident/713946#046bf556036ff58fac4c89be99bad897650ca6436bc27447e9ce0db34a497ed9 Maintenance we are fully upgrading the server because it will explode soon blah blah to serve you better yall dont have school bro but like if needed just tele us or email with (some) of our hearts, Mark Joshwel <mark@joshwel.co> Ryan Lin <ryan@scorchthinks.dev> the server is down, and we knew https://status.joshwel.co/incident/608345 Tue, 24 Jun 2025 15:17:00 -0000 https://status.joshwel.co/incident/608345#e5b4faea303cf79b4d607e37e7c2756a48f32bb88492259632152dd41bacf7ec Incident an unexpected router restart had messed with port forwarding, after a twiddle around router settings (turning port forwarding on and off), we're back up. the server is down, and we know https://status.joshwel.co/incident/602144 Fri, 13 Jun 2025 07:48:00 -0000 https://status.joshwel.co/incident/602144#c5cbf2cc5381c347ad753195eb60110dac16953054adab1eb1457262e56514c2 Incident it was a bootloop, server restored around 1230h SGT the server is down, and we know https://status.joshwel.co/incident/602144 Thu, 12 Jun 2025 22:20:00 -0000 https://status.joshwel.co/incident/602144#48dde79e7bf4350a15e835a9feec37f2cca56bae836503188a5a9e62d10d0597 Incident server is down but ryan is sleeping; it could be a: - heater-related power trip (in which downtime up to six hours may be expected) - or, a router failure (in which downtime up to an hour is expected) the server is down, and we know https://status.joshwel.co/incident/596586 Wed, 04 Jun 2025 01:15:00 -0000 https://status.joshwel.co/incident/596586#ce5494366afe61fcda290bf6837919f6f61524869f3a687081ae92dafd953ce2 Incident aight we up should be good the server is down, and we know https://status.joshwel.co/incident/596586 Wed, 04 Jun 2025 01:09:00 -0000 https://status.joshwel.co/incident/596586#f6b6f1e0d18a6ef19a3aad821b810ad6667249a1db22e4a516614f4f72cd1440 Incident a power trip has caused shit to fly, and services are expected to resume around noon for any qualms contact ryan (at) scorchthinks (dot) dev la poisson du forge is down, and we know https://status.joshwel.co/incident/528981 Tue, 18 Mar 2025 12:31:00 -0000 https://status.joshwel.co/incident/528981#9ba224b4c163a6a04ad960186fb6bbbdcca5bff06e64518a1585fd106058804c Incident the server restoration post-power fault did not properly start the backend before the frontend. resulting in service disruption although the status monitor remained up. the monitored url has been updated and the container stack was fully restarted in proper. no data was lost. la poisson du forge is down, and we know https://status.joshwel.co/incident/528981 Sun, 16 Mar 2025 05:42:00 -0000 https://status.joshwel.co/incident/528981#15ef784d9b86e4f825a25f8c9eec24c618f6cad07f8694ab0d9b5184e0acf91c Incident due to a server power fault, the forge's data is currently inaccessible. an investigation will be launched in the following days. any requests for data can be sent to ryan (at) scorchthinks (dot) dev