#python #ollama #python
The Ollama Python library makes it easy to use Ollama models in your Python projects. To use it, you need to have Ollama installed and running, and then install the library with `pip install ollama`. You can then use simple code to ask questions or generate text using different models. For example, you can ask "Why is the sky blue?" using a specific model like `llama3.2`. The library also supports streaming responses and asynchronous requests, which can be useful for real-time applications. This makes it beneficial for users who want to integrate AI capabilities into their projects quickly and efficiently.
https://github.com/ollama/ollama-python
The Ollama Python library makes it easy to use Ollama models in your Python projects. To use it, you need to have Ollama installed and running, and then install the library with `pip install ollama`. You can then use simple code to ask questions or generate text using different models. For example, you can ask "Why is the sky blue?" using a specific model like `llama3.2`. The library also supports streaming responses and asynchronous requests, which can be useful for real-time applications. This makes it beneficial for users who want to integrate AI capabilities into their projects quickly and efficiently.
https://github.com/ollama/ollama-python
GitHub
GitHub - ollama/ollama-python: Ollama Python library
Ollama Python library. Contribute to ollama/ollama-python development by creating an account on GitHub.
#python #agent #ai #chatbot #chatgpt #docker #function_calling #gemini #gpt #llama #llm #ollama #openai #python #qq #qqbot #qqchannel #telegram
AstrBot is a powerful chatbot and development framework that supports multiple messaging platforms like QQ, WeChat, Telegram, and more. It integrates with large language models (LLMs) such as OpenAI, Google Gemini, and others, allowing for multi-round conversations, personality settings, and multimodal capabilities like image understanding and speech-to-text. The bot has a user-friendly plugin system, a visual management panel, and high stability due to its modular design. This makes it easy to deploy and manage, with various deployment options including Docker, Windows, and Replit. Using AstrBot benefits users by providing a versatile and highly customizable chatbot solution that can be easily extended with new features through plugins.
https://github.com/Soulter/AstrBot
AstrBot is a powerful chatbot and development framework that supports multiple messaging platforms like QQ, WeChat, Telegram, and more. It integrates with large language models (LLMs) such as OpenAI, Google Gemini, and others, allowing for multi-round conversations, personality settings, and multimodal capabilities like image understanding and speech-to-text. The bot has a user-friendly plugin system, a visual management panel, and high stability due to its modular design. This makes it easy to deploy and manage, with various deployment options including Docker, Windows, and Replit. Using AstrBot benefits users by providing a versatile and highly customizable chatbot solution that can be easily extended with new features through plugins.
https://github.com/Soulter/AstrBot
#javascript #ai #llm #llm_ui #llm_webui #llms #ollama #ollama_webui #open_webui #openai #rag #self_hosted #ui #webui
Open WebUI is a powerful and user-friendly AI platform that you can run entirely offline. It supports various AI models like Ollama and OpenAI, and has a built-in inference engine for advanced chat interactions. You can set it up easily using Docker or Kubernetes, and it offers many features such as granular permissions, responsive design for all devices, Markdown and LaTeX support, hands-free voice/video calls, and integration with web searches and image generation.
Using Open WebUI benefits you by providing a secure and customizable AI experience. You can manage user roles and permissions, use multiple models simultaneously, and even create your own models through the Web UI. It also supports multilingual interactions and continuous updates to improve its functionality. Overall, Open WebUI makes it easy to deploy and use AI in a flexible and secure way.
https://github.com/open-webui/open-webui
Open WebUI is a powerful and user-friendly AI platform that you can run entirely offline. It supports various AI models like Ollama and OpenAI, and has a built-in inference engine for advanced chat interactions. You can set it up easily using Docker or Kubernetes, and it offers many features such as granular permissions, responsive design for all devices, Markdown and LaTeX support, hands-free voice/video calls, and integration with web searches and image generation.
Using Open WebUI benefits you by providing a secure and customizable AI experience. You can manage user roles and permissions, use multiple models simultaneously, and even create your own models through the Web UI. It also supports multilingual interactions and continuous updates to improve its functionality. Overall, Open WebUI makes it easy to deploy and use AI in a flexible and secure way.
https://github.com/open-webui/open-webui
GitHub
GitHub - open-webui/open-webui: User-friendly AI Interface (Supports Ollama, OpenAI API, ...)
User-friendly AI Interface (Supports Ollama, OpenAI API, ...) - open-webui/open-webui
#typescript #chat_application #chrome_extension #localllm #ollama #open_source
Page Assist is a browser extension that lets you use your local AI model from any webpage. It works on browsers like Chrome, Brave, Edge, and Firefox. Here’s how it helps you:
- You can open a sidebar or a new tab to interact with your AI model.
- You can ask questions about the content of the webpage you're on.
- It doesn't collect personal data and stores everything locally.
This makes it easy to get help from your AI model while browsing the internet, all without sharing your data.
https://github.com/n4ze3m/page-assist
Page Assist is a browser extension that lets you use your local AI model from any webpage. It works on browsers like Chrome, Brave, Edge, and Firefox. Here’s how it helps you:
- You can open a sidebar or a new tab to interact with your AI model.
- You can ask questions about the content of the webpage you're on.
- It doesn't collect personal data and stores everything locally.
This makes it easy to get help from your AI model while browsing the internet, all without sharing your data.
https://github.com/n4ze3m/page-assist
GitHub
GitHub - n4ze3m/page-assist: Use your locally running AI models to assist you in your web browsing
Use your locally running AI models to assist you in your web browsing - n4ze3m/page-assist
❤1
#typescript #calclaude #chatgpt #claude #cross_platform #desktop #fe #gemini #gemini_pro #gemini_server #gemini_ultra #gpt_4o #groq #nextjs #ollama #react #tauri #tauri_app #vercel #webui
NextChat is a fast and lightweight AI assistant that supports multiple models like Claude, DeepSeek, GPT4, and Gemini Pro. You can use it on the web, or download desktop apps for Windows, MacOS, and Linux. Here are the key benefits You can deploy it for free with one click on Vercel in under a minute.
- **Privacy** It includes markdown support, responsive design, dark mode, and real-time chat capabilities.
- **Customization** Available in several languages including English, Chinese, Japanese, French, Spanish, and more.
Overall, NextChat provides a versatile and secure way to interact with advanced AI models.
https://github.com/ChatGPTNextWeb/NextChat
NextChat is a fast and lightweight AI assistant that supports multiple models like Claude, DeepSeek, GPT4, and Gemini Pro. You can use it on the web, or download desktop apps for Windows, MacOS, and Linux. Here are the key benefits You can deploy it for free with one click on Vercel in under a minute.
- **Privacy** It includes markdown support, responsive design, dark mode, and real-time chat capabilities.
- **Customization** Available in several languages including English, Chinese, Japanese, French, Spanish, and more.
Overall, NextChat provides a versatile and secure way to interact with advanced AI models.
https://github.com/ChatGPTNextWeb/NextChat
GitHub
GitHub - ChatGPTNextWeb/NextChat: ✨ Zero-config AI chat assistant. No API key needed — sign up and instantly chat with GPT-5, Claude…
✨ Zero-config AI chat assistant. No API key needed — sign up and instantly chat with GPT-5, Claude 4, Gemini 2.5, DeepSeek & 100+ top models. Pay-as-you-go saves you more. Available on Web,...
❤1
#python #agent #ai #deepseek #ollama #omniparser #rpa
autoMate is an AI-powered tool that helps automate repetitive tasks on your computer. It works like a digital assistant, using natural language to understand what you need done, so you don't have to write code. This means you can focus on important work while autoMate handles tasks like data organization and report generation automatically. By automating these tasks, you save time and energy, allowing you to be more creative and productive in your work. Plus, it runs locally, keeping your data safe and private.
https://github.com/yuruotong1/autoMate
autoMate is an AI-powered tool that helps automate repetitive tasks on your computer. It works like a digital assistant, using natural language to understand what you need done, so you don't have to write code. This means you can focus on important work while autoMate handles tasks like data organization and report generation automatically. By automating these tasks, you save time and energy, allowing you to be more creative and productive in your work. Plus, it runs locally, keeping your data safe and private.
https://github.com/yuruotong1/autoMate
GitHub
GitHub - yuruotong1/autoMate: Like Manus, Computer Use Agent(CUA) and Omniparser, we are computer-using agents.AI-driven local…
Like Manus, Computer Use Agent(CUA) and Omniparser, we are computer-using agents.AI-driven local automation assistant that uses natural language to make computers work by themselves - yuruotong1/au...
#python #agent #agentops #agents_sdk #ai #anthropic #autogen #cost_estimation #crewai #evals #evaluation_metrics #groq #langchain #llm #mistral #ollama #openai #openai_agents
AgentOps is a tool that helps developers monitor and improve AI agents. It provides features like session replays, cost management for Large Language Models (LLMs), and security checks to prevent data leaks. This platform allows you to track how your agents perform, interact with users, and use external tools. By using AgentOps, you can quickly identify problems, optimize agent performance, and ensure compliance with safety standards. It integrates well with popular platforms like OpenAI and AutoGen, making it easy to set up and use[1][3][5].
https://github.com/AgentOps-AI/agentops
AgentOps is a tool that helps developers monitor and improve AI agents. It provides features like session replays, cost management for Large Language Models (LLMs), and security checks to prevent data leaks. This platform allows you to track how your agents perform, interact with users, and use external tools. By using AgentOps, you can quickly identify problems, optimize agent performance, and ensure compliance with safety standards. It integrates well with popular platforms like OpenAI and AutoGen, making it easy to set up and use[1][3][5].
https://github.com/AgentOps-AI/agentops
GitHub
GitHub - AgentOps-AI/agentops: Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most…
Python SDK for AI agent monitoring, LLM cost tracking, benchmarking, and more. Integrates with most LLMs and agent frameworks including CrewAI, Agno, OpenAI Agents SDK, Langchain, Autogen, AG2, and...
#typescript #aceternity_ui #agent #agents #ai #chrome_extension #extension #fastapi #glean #langchain #langgraph #nextjs #nextjs15 #notebooklm #notion #ollama #perplexity #python #rag #slack #typescript
SurfSense is a highly customizable AI research tool that helps you organize and search your personal knowledge base. It connects to many external sources like search engines, Slack, Notion, YouTube, and GitHub. You can upload various file types and interact with your saved content using natural language. SurfSense provides cited answers and supports local AI models, making it a powerful tool for research. It's also self-hostable and open-source, allowing you to control your data and customize it as needed. This helps you manage information more efficiently and privately.
https://github.com/MODSetter/SurfSense
SurfSense is a highly customizable AI research tool that helps you organize and search your personal knowledge base. It connects to many external sources like search engines, Slack, Notion, YouTube, and GitHub. You can upload various file types and interact with your saved content using natural language. SurfSense provides cited answers and supports local AI models, making it a powerful tool for research. It's also self-hostable and open-source, allowing you to control your data and customize it as needed. This helps you manage information more efficiently and privately.
https://github.com/MODSetter/SurfSense
GitHub
GitHub - MODSetter/SurfSense: Air gapped, open source NotebookLM alternative. Join our Discord: https://discord.gg/ejRNvftDp9
Air gapped, open source NotebookLM alternative. Join our Discord: https://discord.gg/ejRNvftDp9 - MODSetter/SurfSense
#java #anthropic #chatgpt #chroma #embeddings #gemini #gpt #huggingface #java #langchain #llama #milvus #ollama #onnx #openai #openai_api #pgvector #pinecone #vector_database #weaviate
LangChain4j helps you add powerful AI to your Java applications by making it easy to use Large Language Models (LLMs). It provides a simple way to switch between different LLMs and embedding stores without needing to learn each one's specific API. This means you can easily experiment with different models and tools, making your development process faster and more flexible. LangChain4j also offers many examples and tools to help you build complex AI applications quickly, such as chatbots and retrieval systems. This simplifies the integration of AI into your projects, allowing you to focus on creating better applications.
https://github.com/langchain4j/langchain4j
LangChain4j helps you add powerful AI to your Java applications by making it easy to use Large Language Models (LLMs). It provides a simple way to switch between different LLMs and embedding stores without needing to learn each one's specific API. This means you can easily experiment with different models and tools, making your development process faster and more flexible. LangChain4j also offers many examples and tools to help you build complex AI applications quickly, such as chatbots and retrieval systems. This simplifies the integration of AI into your projects, allowing you to focus on creating better applications.
https://github.com/langchain4j/langchain4j
GitHub
GitHub - langchain4j/langchain4j: LangChain4j is an idiomatic, open-source Java library for building LLM-powered applications on…
LangChain4j is an idiomatic, open-source Java library for building LLM-powered applications on the JVM. It offers a unified API over popular LLM providers and vector stores, and makes implementing ...
#typescript #anki #chatgpt #deepseek #electron #evernote #knowledge_base #local_first #markdown #note_taking #notes_app #notion #obsidian #ocr #ollama #openai #pdf #s3 #self_hosted #webdav
SiYuan is a privacy-first personal knowledge management tool. It allows you to organize your thoughts and notes in a secure way, even offline. You can use features like block-level references, Markdown editing, and mathematical formulas. It also supports AI tools and has apps for Android, iOS, and HarmonyOS. SiYuan is open source and free for most features, making it a great choice for managing your personal knowledge securely.
https://github.com/siyuan-note/siyuan
SiYuan is a privacy-first personal knowledge management tool. It allows you to organize your thoughts and notes in a secure way, even offline. You can use features like block-level references, Markdown editing, and mathematical formulas. It also supports AI tools and has apps for Android, iOS, and HarmonyOS. SiYuan is open source and free for most features, making it a great choice for managing your personal knowledge securely.
https://github.com/siyuan-note/siyuan
GitHub
GitHub - siyuan-note/siyuan: An open-source, privacy-first, self-hosted knowledge workspace where humans and AI agents work together…
An open-source, privacy-first, self-hosted knowledge workspace where humans and AI agents work together 开源、隐私优先、自托管的知识工作空间,让人与智能体在此协作 - siyuan-note/siyuan
#kotlin #agentframework #agentic_ai #agents #ai #aiagentframework #android_ai #anthropic #generative_ai #java #jvm #kotlin #ktor #llm #mcp #ollama #openai #spring
Koog is a Kotlin-based open-source framework that helps you build AI agents fully in Kotlin, making it easy to create smart assistants that can use tools, manage complex tasks, and remember past interactions. It supports multiple AI models like OpenAI and Google, runs on many platforms (JVM, JavaScript, iOS), and offers features like real-time streaming, custom tools, and efficient memory use. Koog also provides debugging tools, flexible workflows, and scales from simple chatbots to enterprise systems. Using Koog lets you develop powerful, maintainable AI agents quickly and naturally within the Kotlin ecosystem, benefiting your projects with speed, flexibility, and strong integration options.
https://github.com/JetBrains/koog
Koog is a Kotlin-based open-source framework that helps you build AI agents fully in Kotlin, making it easy to create smart assistants that can use tools, manage complex tasks, and remember past interactions. It supports multiple AI models like OpenAI and Google, runs on many platforms (JVM, JavaScript, iOS), and offers features like real-time streaming, custom tools, and efficient memory use. Koog also provides debugging tools, flexible workflows, and scales from simple chatbots to enterprise systems. Using Koog lets you develop powerful, maintainable AI agents quickly and naturally within the Kotlin ecosystem, benefiting your projects with speed, flexibility, and strong integration options.
https://github.com/JetBrains/koog
GitHub
GitHub - JetBrains/koog: Koog is a JVM (Java and Kotlin) framework for building predictable, fault-tolerant and enterprise-ready…
Koog is a JVM (Java and Kotlin) framework for building predictable, fault-tolerant and enterprise-ready AI agents across all platforms – from backend services to Android and iOS, JVM, and even in-b...
#csharp #agent #ai #avalonia #chat #claude #deepseek #gpt_oss #grok #llm #mcp #ollama #openai #rag #ui_automation
Everywhere is an AI assistant that works directly on your screen without needing screenshots or app switching. You just press a shortcut and it understands the context instantly to help you with tasks like fixing errors, summarizing articles, translating text, or improving your writing tone. It supports many AI models and runs on Windows, with macOS and Linux versions coming soon. This tool saves you time and effort by giving quick, relevant help exactly where you need it, making your work and browsing smoother and more efficient. It also supports multiple languages and has a modern, easy-to-use interface.
https://github.com/DearVa/Everywhere
Everywhere is an AI assistant that works directly on your screen without needing screenshots or app switching. You just press a shortcut and it understands the context instantly to help you with tasks like fixing errors, summarizing articles, translating text, or improving your writing tone. It supports many AI models and runs on Windows, with macOS and Linux versions coming soon. This tool saves you time and effort by giving quick, relevant help exactly where you need it, making your work and browsing smoother and more efficient. It also supports multiple languages and has a modern, easy-to-use interface.
https://github.com/DearVa/Everywhere
GitHub
GitHub - Sylinko/Everywhere: On-screen aware AI assistant for your desktop. Uses current app context, multiple LLMs, and MCP tools…
On-screen aware AI assistant for your desktop. Uses current app context, multiple LLMs, and MCP tools to help you act across apps. - Sylinko/Everywhere
#go #agent #agentic #ai #chatbot #chatbots #embeddings #evaluation #generative_ai #golang #knowledge_base #llm #multi_tenant #multimodel #ollama #openai #question_answering #rag #reranking #semantic_search #vector_search
WeKnora is a powerful tool that helps you understand and find answers in complex documents like PDFs and Word files. It uses advanced AI to read documents, understand what they mean, and answer your questions in a simple way. This tool is useful for businesses and researchers because it can quickly find information from many documents, making it easier to manage knowledge and make decisions. It also supports multiple languages and can be used privately, ensuring your data stays safe.
https://github.com/Tencent/WeKnora
WeKnora is a powerful tool that helps you understand and find answers in complex documents like PDFs and Word files. It uses advanced AI to read documents, understand what they mean, and answer your questions in a simple way. This tool is useful for businesses and researchers because it can quickly find information from many documents, making it easier to manage knowledge and make decisions. It also supports multiple languages and can be used privately, ensuring your data stays safe.
https://github.com/Tencent/WeKnora
GitHub
GitHub - Tencent/WeKnora: Open-source LLM knowledge platform: turn raw documents into a queryable RAG, an autonomous reasoning…
Open-source LLM knowledge platform: turn raw documents into a queryable RAG, an autonomous reasoning agent, and a self-maintaining Wiki. - Tencent/WeKnora
#python #ai #faiss #gpt_oss #langchain #llama_index #llm #localstorage #offline_first #ollama #privacy #python #rag #retrieval_augmented_generation #vector_database #vector_search #vectors
LEANN is a tiny, powerful vector database that lets you turn your laptop into a personal AI assistant capable of searching millions of documents using 97% less storage than traditional systems without losing accuracy. It works by storing a compact graph and computing embeddings only when needed, saving huge space and keeping your data private on your device. You can search your files, emails, browser history, chat logs, live data from platforms like Slack and Twitter, and even codebases—all locally without cloud costs. This means fast, private, and efficient AI-powered search and retrieval on your own laptop.
https://github.com/yichuan-w/LEANN
LEANN is a tiny, powerful vector database that lets you turn your laptop into a personal AI assistant capable of searching millions of documents using 97% less storage than traditional systems without losing accuracy. It works by storing a compact graph and computing embeddings only when needed, saving huge space and keeping your data private on your device. You can search your files, emails, browser history, chat logs, live data from platforms like Slack and Twitter, and even codebases—all locally without cloud costs. This means fast, private, and efficient AI-powered search and retrieval on your own laptop.
https://github.com/yichuan-w/LEANN
GitHub
GitHub - StarTrail-org/LEANN: [MLsys2026 Best Paper]: https://arxiv.org/abs/2506.08276. RAG on Everything with LEANN. Enjoy 97%…
[MLsys2026 Best Paper]: https://arxiv.org/abs/2506.08276. RAG on Everything with LEANN. Enjoy 97% storage savings while running a fast, accurate, and 100% private RAG application on your personal d...
👍1
#python #academia #anthropic #arxiv #brave #deep_research #encryption #home_automation #homeserver #local #local_deep_research #local_llm #mistral #ollama #openai #pubmed #research #research_tool #retrieval_augmented_generation #searxng #self_hosted
Local Deep Research is a free, open-source AI tool you run locally for private, deep research on any topic. It auto-searches the web, academic papers (arXiv, PubMed), and your documents using LLMs like Ollama or GPT, then synthesizes cited reports in minutes. Install easily via Docker or pip, build an encrypted knowledge base from downloads, and get 95% accuracy. Benefits: total privacy (no tracking), zero cost for local models, customizable strategies, and compounding knowledge—saving hours on complex queries while owning your data.
https://github.com/LearningCircuit/local-deep-research
Local Deep Research is a free, open-source AI tool you run locally for private, deep research on any topic. It auto-searches the web, academic papers (arXiv, PubMed), and your documents using LLMs like Ollama or GPT, then synthesizes cited reports in minutes. Install easily via Docker or pip, build an encrypted knowledge base from downloads, and get 95% accuracy. Benefits: total privacy (no tracking), zero cost for local models, customizable strategies, and compounding knowledge—saving hours on complex queries while owning your data.
https://github.com/LearningCircuit/local-deep-research
GitHub
GitHub - LearningCircuit/local-deep-research: ~95% on SimpleQA (e.g. Qwen3.6-27B on a 3090). Supports all local and cloud LLMs…
~95% on SimpleQA (e.g. Qwen3.6-27B on a 3090). Supports all local and cloud LLMs (llama.cpp, Ollama, Google, ...). 10+ search engines - arXiv, PubMed, your private documents. Everything Local &...
#python #agentic_ai #agentic_workflow #agents #function_calling #llama_cpp #llamafile #llm #ollama #python #self_hosted #tool_calling
Forge is a Python tool that makes self-hosted LLM tool-calling more reliable. It helps local models handle multi-step tasks with guardrails, better context control, and support for Ollama, llama-server, Llamafile, and Anthropic. You can use it as a workflow runner, middleware, or proxy server with OpenAI-style clients. The benefit is fewer broken tool calls, better results on small models, and easier setup for agent apps, chat tools, and long-running sessions.
https://github.com/antoinezambelli/forge
Forge is a Python tool that makes self-hosted LLM tool-calling more reliable. It helps local models handle multi-step tasks with guardrails, better context control, and support for Ollama, llama-server, Llamafile, and Anthropic. You can use it as a workflow runner, middleware, or proxy server with OpenAI-style clients. The benefit is fewer broken tool calls, better results on small models, and easier setup for agent apps, chat tools, and long-running sessions.
https://github.com/antoinezambelli/forge
GitHub
GitHub - antoinezambelli/forge: A Python framework for self-hosted LLM tool-calling and multi-step agentic workflows
A Python framework for self-hosted LLM tool-calling and multi-step agentic workflows - antoinezambelli/forge
#python #ai #ai_companion #ai_vtuber #ai_waifu #chatbots #live2d #live2d_web #llm #neuro_sama #ollama
Open-LLM-VTuber is an offline AI companion with voice chat, vision, and a Live2D avatar that works on Windows, macOS, and Linux. It can run as a web app or desktop pet, supports many LLM, speech-to-text, and text-to-speech options, and is highly customizable, so you can build a personal assistant that fits your style and keep chats stored for later.
https://github.com/Open-LLM-VTuber/Open-LLM-VTuber
Open-LLM-VTuber is an offline AI companion with voice chat, vision, and a Live2D avatar that works on Windows, macOS, and Linux. It can run as a web app or desktop pet, supports many LLM, speech-to-text, and text-to-speech options, and is highly customizable, so you can build a personal assistant that fits your style and keep chats stored for later.
https://github.com/Open-LLM-VTuber/Open-LLM-VTuber
GitHub
GitHub - Open-LLM-VTuber/Open-LLM-VTuber: Talk to any LLM with hands-free voice interaction, voice interruption, and Live2D avatar…
Talk to any LLM with hands-free voice interaction, voice interruption, and Live2D avatar running locally across platforms - Open-LLM-VTuber/Open-LLM-VTuber
#python #ai #apple_silicon #benchmarks #cli #command_line_tool #gguf #gpu #huggingface #inference #llm #local_llm #ollama #python #vram
whichllm helps you find the best local AI model that fits your computer, not just the biggest one that fits. It checks your GPU, CPU, and RAM, ranks models using real benchmark data, and can also simulate other hardware, plan upgrades, and start a chat with a model in one command. This saves you time and helps you pick a model that should run well and work better for your needs.
https://github.com/Andyyyy64/whichllm
whichllm helps you find the best local AI model that fits your computer, not just the biggest one that fits. It checks your GPU, CPU, and RAM, ranks models using real benchmark data, and can also simulate other hardware, plan upgrades, and start a chat with a model in one command. This saves you time and helps you pick a model that should run well and work better for your needs.
https://github.com/Andyyyy64/whichllm
GitHub
GitHub - Andyyyy64/whichllm: Find the local LLM that actually runs and performs best on your hardware. Ranked by real, recency…
Find the local LLM that actually runs and performs best on your hardware. Ranked by real, recency-aware benchmarks, not parameter count. One command, run it instantly. - Andyyyy64/whichllm
👍1
#python #cli #consumer_gpu #dpo #fine_tuning #gguf #huggingface #llm #llmops #local_ai #local_llm #lora #low_vram #machine_learning #ollama #peft #python #pytorch #qlora #sft #transformers
Soup is a tool that helps you fine-tune and post-train large language models with one command, using one config file and little setup. It can work on a local GPU, even a 4 GB laptop GPU for an 8B model with layer streaming, so you can train without SSH or cloud hassle. This saves time, reduces setup pain, and makes model training easier to start and manage.
https://github.com/MakazhanAlpamys/Soup
Soup is a tool that helps you fine-tune and post-train large language models with one command, using one config file and little setup. It can work on a local GPU, even a 4 GB laptop GPU for an 8B model with layer streaming, so you can train without SSH or cloud hassle. This saves time, reduces setup pain, and makes model training easier to start and manage.
https://github.com/MakazhanAlpamys/Soup
GitHub
GitHub - MakazhanAlpamys/Soup: Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU.
Fine-tune LLMs from one YAML. Layer streaming trains an 8B model on a 4 GB laptop GPU. - MakazhanAlpamys/Soup
❤1