<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/">
  <channel>
    <title>AI on Osman&#39;s Odyssey: Byte &amp; Build</title>
    <link>https://www.ahmadosman.com/tags/ai/</link>
    <description>Recent content in AI on Osman&#39;s Odyssey: Byte &amp; Build</description>
    <image>
      <title>Osman&#39;s Odyssey: Byte &amp; Build</title>
      <url>https://www.ahmadosman.com/logo/byte-and-build.png</url>
      <link>https://www.ahmadosman.com/logo/byte-and-build.png</link>
    </image>
    <generator>Hugo -- 0.145.0</generator>
    <language>en-us</language>
    <lastBuildDate>Sat, 21 Jun 2025 07:08:00 -0500</lastBuildDate>
    <atom:link href="https://www.ahmadosman.com/tags/ai/index.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Software Engineers Aren&#39;t Getting Automated—Local AI Has To Win</title>
      <link>https://www.ahmadosman.com/blog/software-engineers-arent-getting-automated-local-ai-has-to-win/</link>
      <pubDate>Sat, 21 Jun 2025 07:08:00 -0500</pubDate>
      <guid>https://www.ahmadosman.com/blog/software-engineers-arent-getting-automated-local-ai-has-to-win/</guid>
      <description>Stop worrying about AI replacing you—the real threat is losing technical depth. As cloud dependence grows and platforms get more opaque, local-first AI, open weights, and full-stack ownership are the only safety nets left. Why the future belongs to those who can build, debug, and own their tools from the metal up. Trust no corporate overlord.</description>
    </item>
    <item>
      <title>Just Like GPUs, We Need To Be Stress Tested</title>
      <link>https://www.ahmadosman.com/blog/just-like-gpus-we-need-to-be-stress-tested-101-days-of-blogging/</link>
      <pubDate>Wed, 18 Jun 2025 23:02:02 -0500</pubDate>
      <guid>https://www.ahmadosman.com/blog/just-like-gpus-we-need-to-be-stress-tested-101-days-of-blogging/</guid>
      <description>Why 101 days of daily tech blogging? A raw, open challenge on AI, LLMs, self-hosted experiments, knowledge distillation, and why consistency beats talent. Expect rants, technical breakdowns, open hardware journeys, memes, and daily accountability from the basement AI server guy.</description>
    </item>
    <item>
      <title>Once Undesirable, Now Undeniable</title>
      <link>https://www.ahmadosman.com/blog/once-undesirable-now-undeniable/</link>
      <pubDate>Wed, 30 Apr 2025 15:46:46 -0500</pubDate>
      <guid>https://www.ahmadosman.com/blog/once-undesirable-now-undeniable/</guid>
      <description>How taking risks, building in public, and refusing to play a losing game flipped the script—and why sometimes you have to become undeniable before you ever become accepted.</description>
    </item>
    <item>
      <title>Build Your Private AI Screenshot Organizer with LMStudio</title>
      <link>https://www.ahmadosman.com/blog/build-your-local-privat-ai-screenshot-organizer-with-lmstudio/</link>
      <pubDate>Tue, 22 Apr 2025 08:56:56 -0500</pubDate>
      <guid>https://www.ahmadosman.com/blog/build-your-local-privat-ai-screenshot-organizer-with-lmstudio/</guid>
      <description>Build a local, privacy-first screenshot organizer using LMStudio’s Python SDK and Gemma 3 multimodal models. Keep your data off the cloud, automate screenshot categorization, and leverage the power of open-source AI—all running from your own PC. Step-by-step guide, code walkthrough, and a practical use-case for local LLMs.</description>
    </item>
    <item>
      <title>Stop Wasting Your Multi-GPU Setup With llama.cpp</title>
      <link>https://www.ahmadosman.com/blog/do-not-use-llama-cpp-or-ollama-on-multi-gpus-setups-use-vllm-or-exllamav2/</link>
      <pubDate>Fri, 07 Feb 2025 05:06:36 -0600</pubDate>
      <guid>https://www.ahmadosman.com/blog/do-not-use-llama-cpp-or-ollama-on-multi-gpus-setups-use-vllm-or-exllamav2/</guid>
      <description>Exploring the intricacies of Inference Engines and why llama.cpp should be avoided when running Multi-GPU setups. Learn about Tensor Parallelism, the role of vLLM in batch inference, and why ExLlamaV2 has been a game-changer for GPU-optimized AI serving since it introduced Tensor Parallelism.</description>
    </item>
    <item>
      <title>Resources From X/Twitter Audio Space on LLMs &amp; AI - 2025-02-02</title>
      <link>https://www.ahmadosman.com/blog/deepseek-r1-space/</link>
      <pubDate>Sun, 02 Feb 2025 15:35:35 -0600</pubDate>
      <guid>https://www.ahmadosman.com/blog/deepseek-r1-space/</guid>
      <description>A curated collection of links, books, tools, and benchmarks discussed during the February 2nd, 2025 Twitter/X Audio Space on LLMs and AI. Includes practical resources, RAG leaderboards, toolkits, and perspectives on AI adoption in the Middle East and globally.</description>
    </item>
    <item>
      <title>Antifragile AI</title>
      <link>https://www.ahmadosman.com/blog/taleb-antifragile-ai-insights/</link>
      <pubDate>Tue, 03 Dec 2024 04:21:55 -0600</pubDate>
      <guid>https://www.ahmadosman.com/blog/taleb-antifragile-ai-insights/</guid>
      <description>Explore how AI systems can become antifragile, harnessing uncertainty to thrive. Learn about the shift and acceleration from traditional software to AI agentic systems and their implications for the future.</description>
    </item>
    <item>
      <title>Serving AI From The Basement — Part II</title>
      <link>https://www.ahmadosman.com/blog/serving-ai-from-the-basement-part-ii/</link>
      <pubDate>Wed, 18 Sep 2024 05:57:26 -0500</pubDate>
      <guid>https://www.ahmadosman.com/blog/serving-ai-from-the-basement-part-ii/</guid>
      <description>SWE Agentic Framework, MoEs, Quantizations &amp; Mixed Precision, Batch Inference, LLM Architectures, vLLM, DeepSeek v2.5, Embedding Models, and Speculative Decoding: An LLM Brain Dump... I have been working on a multi-agent system that simulates a team of Software Engineers; this system assigns projects, creates teams and adds members to them based on areas of expertise and need, and asks team members to build features, assign story points, have pair programming sessions together, etc.</description>
    </item>
    <item>
      <title>Serving AI From The Basement — Part I</title>
      <link>https://www.ahmadosman.com/blog/serving-ai-from-the-basement-part-i/</link>
      <pubDate>Fri, 06 Sep 2024 16:37:23 -0500</pubDate>
      <guid>https://www.ahmadosman.com/blog/serving-ai-from-the-basement-part-i/</guid>
      <description>Dedicated LLM server powered by 8x RTX 3090 Graphic Cards, boasting a total of 192GB of VRAM.</description>
    </item>
  </channel>
</rss>
