Found 564 bookmarks
Newest
Running LLaMA Locally with Llama.cpp: A Complete Guide
Running LLaMA Locally with Llama.cpp: A Complete Guide
Llama.cpp is a powerful and efficient inference framework for running LLaMA models locally on your machine. Unlike other tools such as…
·medium.com·
Running LLaMA Locally with Llama.cpp: A Complete Guide
The Complete Developer's Guide to Running LLMs Locally
The Complete Developer's Guide to Running LLMs Locally
A comprehensive guide covering the local LLM stack from hardware requirements to production deployment. Compare Ollama, LM Studio, llama.cpp and build your first local AI application.
·sitepoint.com·
The Complete Developer's Guide to Running LLMs Locally
LM Studio Developer Docs | LM Studio Docs
LM Studio Developer Docs | LM Studio Docs
Build with LM Studio's local APIs and SDKs — TypeScript, Python, REST, and OpenAI and Anthropic-compatible endpoints.
·lmstudio.ai·
LM Studio Developer Docs | LM Studio Docs
Your GPUs Just Got 6x More Valuable. No New Hardware Required.
Your GPUs Just Got 6x More Valuable. No New Hardware Required.
Watch now | The variable that decides who wins the AI infrastructure war isn’t a faster chip or a better model. It’s a compression algorithm.
·natesnewsletter.substack.com·
Your GPUs Just Got 6x More Valuable. No New Hardware Required.
Who’s the Admin, Me or Claude?
Who’s the Admin, Me or Claude?
Credit: Museums Victoria / Unsplash There’s a lot of conversation right now about “context engineering” for dev work; structuring what you feed an LLM so it can do useful things. …
·cate.blog·
Who’s the Admin, Me or Claude?
Mastering Caching Methods in Large Language Models (LLMs)
Mastering Caching Methods in Large Language Models (LLMs)
Large Language Models (LLMs) like OpenAI’s GPT-4 have transformed natural language processing, enabling applications ranging from chatbots…
·masteringllm.medium.com·
Mastering Caching Methods in Large Language Models (LLMs)
How to Implement Effective LLM Caching
How to Implement Effective LLM Caching
A deep dive into effective caching strategies for building scalable and cost-efficient LLM applications, covering exact key vs. semantic caching, architectural patterns, and practical implementation tips.
·helicone.ai·
How to Implement Effective LLM Caching
Anatomy of the .claude/ Folder
Anatomy of the .claude/ Folder
A complete guide to CLAUDE.md, custom commands, skills, agents, and permissions, and how to set them up properly.
·blog.dailydoseofds.com·
Anatomy of the .claude/ Folder
Harper's Policy on Agent PRs
Harper's Policy on Agent PRs
The goal of this page is to formalize my answer so that we can judiciously deal with patch requests produced by LLMs.
·elijahpotter.dev·
Harper's Policy on Agent PRs
The 8 Levels of Agentic Engineering — Bassim Eledath
The 8 Levels of Agentic Engineering — Bassim Eledath
AI's coding ability is outpacing our ability to wield it effectively. That gap closes in levels — 8 of them. Here's the progression from tab complete to autonomous agent teams.
·bassimeledath.com·
The 8 Levels of Agentic Engineering — Bassim Eledath
Top AI coding tools make mistakes one in four times, study shows
Top AI coding tools make mistakes one in four times, study shows
New research from the University of Waterloo shows that artificial intelligence (AI) still struggles with some basic software development tasks, raising questions about how reliably AI systems can assist ...
·techxplore.com·
Top AI coding tools make mistakes one in four times, study shows
Claude Skill incoming! Generating Postman collections with AI
Claude Skill incoming! Generating Postman collections with AI
When speed matters more than perfection, API documentation can quickly become a bottleneck. In this post, I share how we used thoughtbot’s Claude Skill to generate Postman collections directly from a Rails codebase.
·thoughtbot.com·
Claude Skill incoming! Generating Postman collections with AI