docs: describe the sglang endpoint behind caddy
Assisted-by: GLM 5.3 Flash
This commit is contained in:
@@ -22,7 +22,7 @@ own builtins, and none of them needs a module installed beside it.
|
||||
automatic updates); `workstation-setup.pl` sets up a Fedora desktop (Brave, the
|
||||
official Go toolchain, Rust, GoLand, Flatpak applications, firewall, SELinux).
|
||||
- **Deployment**: `sglang-deploy.pl` puts an SGLang inference server on an AMD GPU
|
||||
behind nginx with HTTPS and an API key, runs the engine from the project's ROCm
|
||||
behind Caddy with HTTPS and an API key, runs the engine from the project's ROCm
|
||||
container image, and downloads the model weights from ModelScope.
|
||||
- **Maintenance**: `system-optimise.pl` removes old kernels (the running one and one
|
||||
fallback always stay), vacuums journals, clears temporary files and core dumps, and
|
||||
@@ -92,7 +92,7 @@ a change would do without doing it.
|
||||
| `security-audit.pl` | 2.0.0 | Security posture with a grade from A to F and an exit code for cron. Linux only |
|
||||
| `server-setup.pl` | 2.0.0 | Server initial setup for Fedora, CentOS Stream and openEuler |
|
||||
| `workstation-setup.pl` | 2.0.0 | Fedora desktop setup |
|
||||
| `sglang-deploy.pl` | 2.0.0 | SGLang inference server on an AMD GPU, behind nginx with HTTPS |
|
||||
| `sglang-deploy.pl` | 2.1.0 | SGLang inference server on an AMD GPU, behind Caddy with HTTPS |
|
||||
| `system-optimise.pl` | 2.0.0 | System cleanup; refuses rpm-ostree systems |
|
||||
|
||||
The full flag reference is in [docs/CLI.md](docs/CLI.md).
|
||||
|
||||
Reference in New Issue
Block a user