3 ms·
Airgapped LLM inferrence server can't serve their output tokens, right?
by tintor 13d ago
Airgapped LLM inferrence server can't serve their output tokens, right?
- angry_octet 13d agoThey can expose just their inference port, possible via some supervisor. The inference consumer can also be air gapped. This kind of segmentation is increasingly common for high value services.
- what 13d agoThen it’s not air gapped…
- angry_octet 12d agoIt's air gapped from e.g. the internet.
- bibimsz 13d agonot at a high bitrate