
Running local LLMs on your own hardware is no longer the domain of the chosen few with H100 clusters. Or is it still? 😁
🗓 Today at 6:00 PM Moscow time there will be a live broadcast in «Assembly Point». We will be discussing the installation and use of local models.
What will be covered:
▪️ How to choose a suitable model for your hardware so that it doesn't consume all system memory and works with adequate TPS (tokens per second).
▪️ Setting up the model in chat mode for everyday tasks.
▪️ Connecting a local LLM to an agentic IDE (using Kilo Code as an example).
▪️ Routing requests to the local model via LangChain.
▪️ Is it even necessary?
The broadcast and recording will be available to participants of «Assembly Point». Access is arranged through the bot: https://t.me/TScompiler_bot
Comments
0No comments yet.
Sign in to join the discussion.