Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

Is there any way to get abortable streaming responses from Llama 2 (whether from Replicate or elsewhere) in the way you currently can using ChatGPT?


KoboldCPP or text-gen-ui




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: