{"id":7077,"date":"2026-10-07T07:23:51","date_gmt":"2026-10-07T07:23:51","guid":{"rendered":"https:\/\/areeblog.com\/?p=7077"},"modified":"2026-10-07T07:28:15","modified_gmt":"2026-10-07T07:28:15","slug":"opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference","status":"publish","type":"post","link":"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/","title":{"rendered":"OpenTPU Tests Whether AI Agents Can Design the Hardware That Runs AI Inference"},"content":{"rendered":"<p><img loading=\"lazy\" loading=\"lazy\" decoding=\"async\" data-attachment-id=\"7080\" data-permalink=\"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/ai-platform-blog-chat-labs-featured-2\/\" data-orig-file=\"https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1.png\" data-orig-size=\"1920,1080\" data-comments-opened=\"1\" data-image-meta=\"{&quot;aperture&quot;:&quot;0&quot;,&quot;credit&quot;:&quot;&quot;,&quot;camera&quot;:&quot;&quot;,&quot;caption&quot;:&quot;&quot;,&quot;created_timestamp&quot;:&quot;0&quot;,&quot;copyright&quot;:&quot;&quot;,&quot;focal_length&quot;:&quot;0&quot;,&quot;iso&quot;:&quot;0&quot;,&quot;shutter_speed&quot;:&quot;0&quot;,&quot;title&quot;:&quot;&quot;,&quot;orientation&quot;:&quot;0&quot;,&quot;alt&quot;:&quot;&quot;}\" data-image-title=\"ai-platform-blog-chat-labs-featured\" data-image-description=\"\" data-image-caption=\"\" data-large-file=\"https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1-1024x576.png\" class=\"aligncenter size-full wp-image-7080\" src=\"https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1.png\" alt=\"OpenTPU Tests Whether AI Agents Can Design the Hardware That Runs AI Inference\" width=\"1920\" height=\"1080\" srcset=\"https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1.png 1920w, https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1-300x169.png 300w, https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1-1024x576.png 1024w, https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1-768x432.png 768w, https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1-1536x864.png 1536w, https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1-860x484.png 860w\" sizes=\"auto, (max-width: 1920px) 100vw, 1920px\" \/><\/p>\n<p>An open-source project called <strong>openTPU<\/strong> is testing whether <a href=\"https:\/\/areeblog.com\/docker-and-gitguardian-add-secret-scanning-to-ai-coding-sandboxes\/\">AI coding<\/a> agents can contribute to the design of the hardware used to run AI models, with a working accelerator now executing modern language models on an inexpensive Xilinx Kintex-7 FPGA board.<\/p>\n<p>Developed by Felipe Sens Bonetto, <a href=\"https:\/\/github.com\/FeSens\/openTPU\">the openTPU repository<\/a> brings the accelerator&#8217;s hardware and software stack into a single project. It includes the SystemVerilog RTL, a custom instruction set, an instruction-level simulator, a kernel language, compiler, host software and FPGA build files.<\/p>\n<p>The project describes its central question in more direct terms: how far AI agents can go in hardware design, and whether they can build the hardware that runs their own inference. The current implementation is an FPGA accelerator rather than a fabricated AI chip, but it has been tested on physical hardware with real model weights.<\/p>\n<h2>A complete inference stack on one FPGA<\/h2>\n<p>The openTPU hardware targets an <strong>Inspur YPCB-00338<\/strong> accelerator card built around a <strong>Xilinx Kintex-7 XC7K480T-FFG1156-2<\/strong> FPGA. The board provides two 2 GiB DDR3 memory channels for 4 GiB of total local memory and connects to the host system through PCIe.<\/p>\n<p>The accelerator&#8217;s production configuration uses a 128-wide data dimension, four matrix-unit columns and eight vector-processing lanes. It runs at 133.33 MHz and has a theoretical combined DDR3 bandwidth of about 17.1 GB\/s.<\/p>\n<p>Rather than relying on a large software stack outside the project, openTPU exposes the path from model kernels to hardware instructions. Its compiler turns kernels written in the project&#8217;s small kernel language into the accelerator&#8217;s instruction set, while the simulator provides an executable reference for checking the hardware&#8217;s behavior.<\/p>\n<p>The repository also includes host software and utilities for controlling and profiling the accelerator. The design does not use a cache or a hidden scheduler; memory movement is explicitly represented in the program.<\/p>\n<h2>Qwen, Gemma and LFM2.5 run on the board<\/h2>\n<p>The project has moved beyond simulation and demonstrated inference using actual model weights on the Kintex-7 card. Its published benchmark set covers models including <strong>Qwen3, Qwen3.5, LFM2.5 and Gemma 4<\/strong>, along with several larger models.<\/p>\n<p>For example, the project reports a device decode rate of <strong>85.8 tokens per second<\/strong> for the 4-bit version of LFM2.5-230M, compared with 59.0 tokens per second using int8 weights. The 4-bit Qwen3-0.6B configuration reaches 31.3 tokens per second, while Qwen3.5-0.8B reaches 24.5 tokens per second.<\/p>\n<p>Gemma 4 E2B is reported at 12.14 tokens per second using the project&#8217;s 4-bit configuration. The benchmark measurements were carried out on the physical accelerator rather than being estimates derived only from simulation.<\/p>\n<p>The reported figures distinguish accelerator execution from complete host-to-device wall-clock performance. In the project&#8217;s benchmark, decoding uses 64 greedy tokens after a 512-token prompt, with some host-side work included in the wall-time measurement.<\/p>\n<p>The project also reports models that exceed the board&#8217;s 4 GiB local memory capacity. For mixture-of-experts models, it can keep selected experts on the FPGA and stream others from host storage over PCIe.<\/p>\n<p>That approach has been demonstrated with <strong>LFM2.5-8B-A1B<\/strong>, which the project reports at 10.6 tokens per second, and <strong>Qwen3.5-35B-A3B<\/strong>, reported at 3.95 tokens per second. The larger model has 34.7 billion total parameters but about 3.0 billion active parameters, according to the project&#8217;s measurements.<\/p>\n<h2>AI-generated hardware is being tested against physical constraints<\/h2>\n<p>The question behind openTPU is more demanding than asking an AI model to generate SystemVerilog that compiles. Hardware changes have to survive simulation, synthesis, implementation and timing constraints before they can be useful on a real FPGA.<\/p>\n<p>The project draws on the methodology used in Bonetto&#8217;s related <a href=\"https:\/\/github.com\/FeSens\/auto-arch-tournament\">auto-arch-tournament<\/a> project. That system uses AI coding agents to propose and implement architectural changes, then evaluates each candidate with a fixed pipeline involving hardware checks, benchmarks and implementation tests. A candidate only becomes the new design when it produces a measurable improvement while remaining valid.<\/p>\n<p>That earlier experiment evaluated 73 hardware hypotheses, with 10 winning changes eventually merged. The reported result improved its measured fitness from about 301 iterations per second to roughly 577 iterations per second.<\/p>\n<p>For openTPU, the same general approach shifts the target from a CPU experiment to an AI inference accelerator. The important distinction is that the agents operate within a human-defined development and verification environment. The available evidence does not show an AI system independently conceiving, fabricating and validating a commercial AI chip from scratch.<\/p>\n<p>The project instead demonstrates an AI-assisted hardware engineering loop in which generated changes can be checked against an instruction-level simulator and then taken through real FPGA implementation.<\/p>\n<h2>Performance is limited by the memory system<\/h2>\n<p>The Kintex-7 implementation is also constrained by the hardware it runs on. The project reports that inference decoding is largely <strong>DRAM-bound<\/strong>, because model weights have to be streamed repeatedly for each generated token.<\/p>\n<p>During decoding, the accelerator is reported to use roughly 82% to 94% of the available DRAM bandwidth, depending on the model and configuration. The project identifies memory efficiency and timing margin as continuing engineering constraints.<\/p>\n<p>The production design reportedly closes timing at 133.33 MHz with only about 0.032 ns of worst-negative-slack margin, showing how tightly the implementation is operating within the FPGA&#8217;s limits.<\/p>\n<p>The project&#8217;s 4-bit configurations reduce weight traffic further. Its documentation describes a custom 4-bit representation averaging about 4.25 bits per weight, with the language-model head generally remaining at int8. The trade-off is lower memory traffic and higher decode throughput at the cost of some model-quality degradation.<\/p>\n<p>For the larger models that cannot fit entirely into local memory, PCIe becomes another constraint. The project reports that Qwen3.5-35B-A3B requires substantial expert transfers between host storage and the accelerator during decoding.<\/p>\n<p>openTPU&#8217;s repository states that hardware execution and the project&#8217;s ISA simulator produce matching tokens in its tests, providing a direct software reference for the physical implementation.<\/p>\n","protected":false},"excerpt":{"rendered":"<p>An open-source project called openTPU is testing whether AI coding agents can contribute to the design of the hardware used to run AI models, with a working accelerator now executing modern language models on an inexpensive Xilinx Kintex-7 FPGA board. Developed by Felipe Sens Bonetto, the openTPU repository brings the accelerator&#8217;s hardware and software stack [&hellip;]<\/p>\n","protected":false},"author":2,"featured_media":7080,"comment_status":"open","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"content-type":"","_monsterinsights_skip_tracking":false,"_jetpack_newsletter_access":"","_jetpack_dont_email_post_to_subs":false,"_jetpack_newsletter_tier_id":0,"_jetpack_memberships_contains_paywalled_content":false,"_jetpack_feature_clip_id":0,"_jetpack_memberships_contains_paid_content":false,"footnotes":"","jetpack_post_was_ever_published":false},"categories":[164],"tags":[166],"class_list":["post-7077","post","type-post","status-publish","format-standard","has-post-thumbnail","category-tech-updates","tag-ai"],"share_on_mastodon":{"url":"https:\/\/mastodon.social\/@Areeblog\/117398467111475117","error":""},"yoast_head":"<!-- This site is optimized with the Yoast SEO Premium plugin v28.4 (Yoast SEO v28.6) - https:\/\/yoast.com\/product\/yoast-seo-premium-wordpress\/ -->\n<title>OpenTPU Tests Whether AI Agents Can Design the Hardware That Runs AI Inference - Aree Blog<\/title>\n<meta name=\"description\" content=\"OpenTPU tests AI-designed hardware by running Qwen3, Gemma 4 and LFM2.5 models on a Kintex-7 FPGA accelerator.\" \/>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/\" \/>\n<meta property=\"og:locale\" content=\"en_US\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"OpenTPU Tests Whether AI Agents Can Design the Hardware That Runs AI Inference\" \/>\n<meta property=\"og:description\" content=\"OpenTPU tests AI-designed hardware by running Qwen3, Gemma 4 and LFM2.5 models on a Kintex-7 FPGA accelerator.\" \/>\n<meta property=\"og:url\" content=\"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/\" \/>\n<meta property=\"og:site_name\" content=\"Aree Blog\" \/>\n<meta property=\"article:published_time\" content=\"2026-10-07T07:23:51+00:00\" \/>\n<meta property=\"article:modified_time\" content=\"2026-10-07T07:28:15+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1.png\" \/>\n\t<meta property=\"og:image:width\" content=\"1920\" \/>\n\t<meta property=\"og:image:height\" content=\"1080\" \/>\n\t<meta property=\"og:image:type\" content=\"image\/png\" \/>\n<meta name=\"author\" content=\"Daniel Chinonso John\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:label1\" content=\"Written by\" \/>\n\t<meta name=\"twitter:data1\" content=\"Daniel Chinonso John\" \/>\n\t<meta name=\"twitter:label2\" content=\"Est. reading time\" \/>\n\t<meta name=\"twitter:data2\" content=\"5 minutes\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\\\/\\\/schema.org\",\"@graph\":[{\"@type\":\"Article\",\"@id\":\"https:\\\/\\\/areeblog.com\\\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\\\/#article\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/areeblog.com\\\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\\\/\"},\"author\":{\"name\":\"Daniel Chinonso John\",\"@id\":\"https:\\\/\\\/areeblog.com\\\/#\\\/schema\\\/person\\\/d972222c55618fb0f4b4c0c11ff52f63\"},\"headline\":\"OpenTPU Tests Whether AI Agents Can Design the Hardware That Runs AI Inference\",\"datePublished\":\"2026-10-07T07:23:51+00:00\",\"dateModified\":\"2026-10-07T07:28:15+00:00\",\"mainEntityOfPage\":{\"@id\":\"https:\\\/\\\/areeblog.com\\\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\\\/\"},\"wordCount\":968,\"commentCount\":0,\"image\":{\"@id\":\"https:\\\/\\\/areeblog.com\\\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/areeblog.com\\\/wp-content\\\/uploads\\\/2026\\\/10\\\/ai-platform-blog-chat-labs-featured-1.png\",\"keywords\":[\"AI\"],\"articleSection\":[\"Tech Updates\"],\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"CommentAction\",\"name\":\"Comment\",\"target\":[\"https:\\\/\\\/areeblog.com\\\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\\\/#respond\"]}]},{\"@type\":\"WebPage\",\"@id\":\"https:\\\/\\\/areeblog.com\\\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\\\/\",\"url\":\"https:\\\/\\\/areeblog.com\\\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\\\/\",\"name\":\"OpenTPU Tests Whether AI Agents Can Design the Hardware That Runs AI Inference - Aree Blog\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/areeblog.com\\\/#website\"},\"primaryImageOfPage\":{\"@id\":\"https:\\\/\\\/areeblog.com\\\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\\\/#primaryimage\"},\"image\":{\"@id\":\"https:\\\/\\\/areeblog.com\\\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/areeblog.com\\\/wp-content\\\/uploads\\\/2026\\\/10\\\/ai-platform-blog-chat-labs-featured-1.png\",\"datePublished\":\"2026-10-07T07:23:51+00:00\",\"dateModified\":\"2026-10-07T07:28:15+00:00\",\"author\":{\"@id\":\"https:\\\/\\\/areeblog.com\\\/#\\\/schema\\\/person\\\/d972222c55618fb0f4b4c0c11ff52f63\"},\"description\":\"OpenTPU tests AI-designed hardware by running Qwen3, Gemma 4 and LFM2.5 models on a Kintex-7 FPGA accelerator.\",\"breadcrumb\":{\"@id\":\"https:\\\/\\\/areeblog.com\\\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\\\/#breadcrumb\"},\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\\\/\\\/areeblog.com\\\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\\\/\"]}]},{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/areeblog.com\\\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\\\/#primaryimage\",\"url\":\"https:\\\/\\\/areeblog.com\\\/wp-content\\\/uploads\\\/2026\\\/10\\\/ai-platform-blog-chat-labs-featured-1.png\",\"contentUrl\":\"https:\\\/\\\/areeblog.com\\\/wp-content\\\/uploads\\\/2026\\\/10\\\/ai-platform-blog-chat-labs-featured-1.png\",\"width\":1920,\"height\":1080,\"caption\":\"OpenTPU Tests Whether AI Agents Can Design the Hardware That Runs AI Inference\"},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\\\/\\\/areeblog.com\\\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\\\/#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\\\/\\\/areeblog.com\\\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"OpenTPU Tests Whether AI Agents Can Design the Hardware That Runs AI Inference\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\\\/\\\/areeblog.com\\\/#website\",\"url\":\"https:\\\/\\\/areeblog.com\\\/\",\"name\":\"Aree Blog\",\"description\":\"Unfiltered Perspectives, Unstoppable Insights\",\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\\\/\\\/areeblog.com\\\/?s={search_term_string}\"},\"query-input\":{\"@type\":\"PropertyValueSpecification\",\"valueRequired\":true,\"valueName\":\"search_term_string\"}}],\"inLanguage\":\"en-US\"},{\"@type\":\"Person\",\"@id\":\"https:\\\/\\\/areeblog.com\\\/#\\\/schema\\\/person\\\/d972222c55618fb0f4b4c0c11ff52f63\",\"name\":\"Daniel Chinonso John\",\"description\":\"Daniel Chinonso John is a web designer, penetration tester, and founder of Aree Tech. He writes clear, actionable posts at the intersection of productivity, AI, cybersecurity, and blogging to help readers get things done.\",\"sameAs\":[\"https:\\\/\\\/www.linkedin.com\\\/in\\\/daniel-john-45183a169\\\/\"],\"url\":\"https:\\\/\\\/areeblog.com\\\/author\\\/danojohn55gmail-com\\\/\"}]}<\/script>\n<!-- \/ Yoast SEO Premium plugin. -->","yoast_head_json":{"title":"OpenTPU Tests Whether AI Agents Can Design the Hardware That Runs AI Inference - Aree Blog","description":"OpenTPU tests AI-designed hardware by running Qwen3, Gemma 4 and LFM2.5 models on a Kintex-7 FPGA accelerator.","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/","og_locale":"en_US","og_type":"article","og_title":"OpenTPU Tests Whether AI Agents Can Design the Hardware That Runs AI Inference","og_description":"OpenTPU tests AI-designed hardware by running Qwen3, Gemma 4 and LFM2.5 models on a Kintex-7 FPGA accelerator.","og_url":"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/","og_site_name":"Aree Blog","article_published_time":"2026-10-07T07:23:51+00:00","article_modified_time":"2026-10-07T07:28:15+00:00","og_image":[{"width":1920,"height":1080,"url":"https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1.png","type":"image\/png"}],"author":"Daniel Chinonso John","twitter_card":"summary_large_image","twitter_misc":{"Written by":"Daniel Chinonso John","Est. reading time":"5 minutes"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"Article","@id":"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/#article","isPartOf":{"@id":"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/"},"author":{"name":"Daniel Chinonso John","@id":"https:\/\/areeblog.com\/#\/schema\/person\/d972222c55618fb0f4b4c0c11ff52f63"},"headline":"OpenTPU Tests Whether AI Agents Can Design the Hardware That Runs AI Inference","datePublished":"2026-10-07T07:23:51+00:00","dateModified":"2026-10-07T07:28:15+00:00","mainEntityOfPage":{"@id":"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/"},"wordCount":968,"commentCount":0,"image":{"@id":"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/#primaryimage"},"thumbnailUrl":"https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1.png","keywords":["AI"],"articleSection":["Tech Updates"],"inLanguage":"en-US","potentialAction":[{"@type":"CommentAction","name":"Comment","target":["https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/#respond"]}]},{"@type":"WebPage","@id":"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/","url":"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/","name":"OpenTPU Tests Whether AI Agents Can Design the Hardware That Runs AI Inference - Aree Blog","isPartOf":{"@id":"https:\/\/areeblog.com\/#website"},"primaryImageOfPage":{"@id":"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/#primaryimage"},"image":{"@id":"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/#primaryimage"},"thumbnailUrl":"https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1.png","datePublished":"2026-10-07T07:23:51+00:00","dateModified":"2026-10-07T07:28:15+00:00","author":{"@id":"https:\/\/areeblog.com\/#\/schema\/person\/d972222c55618fb0f4b4c0c11ff52f63"},"description":"OpenTPU tests AI-designed hardware by running Qwen3, Gemma 4 and LFM2.5 models on a Kintex-7 FPGA accelerator.","breadcrumb":{"@id":"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/"]}]},{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/#primaryimage","url":"https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1.png","contentUrl":"https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1.png","width":1920,"height":1080,"caption":"OpenTPU Tests Whether AI Agents Can Design the Hardware That Runs AI Inference"},{"@type":"BreadcrumbList","@id":"https:\/\/areeblog.com\/opentpu-tests-whether-ai-agents-can-design-the-hardware-that-runs-ai-inference\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/areeblog.com\/"},{"@type":"ListItem","position":2,"name":"OpenTPU Tests Whether AI Agents Can Design the Hardware That Runs AI Inference"}]},{"@type":"WebSite","@id":"https:\/\/areeblog.com\/#website","url":"https:\/\/areeblog.com\/","name":"Aree Blog","description":"Unfiltered Perspectives, Unstoppable Insights","potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/areeblog.com\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"en-US"},{"@type":"Person","@id":"https:\/\/areeblog.com\/#\/schema\/person\/d972222c55618fb0f4b4c0c11ff52f63","name":"Daniel Chinonso John","description":"Daniel Chinonso John is a web designer, penetration tester, and founder of Aree Tech. He writes clear, actionable posts at the intersection of productivity, AI, cybersecurity, and blogging to help readers get things done.","sameAs":["https:\/\/www.linkedin.com\/in\/daniel-john-45183a169\/"],"url":"https:\/\/areeblog.com\/author\/danojohn55gmail-com\/"}]}},"jetpack_sharing_enabled":true,"jetpack-related-posts":[{"id":7066,"url":"https:\/\/areeblog.com\/developers-repurpose-crypto-mining-fpgas-to-run-qwen3-5-ai-models\/","url_meta":{"origin":7077,"position":0},"title":"Developers Repurpose Crypto-Mining FPGAs to Run Qwen3.5 AI Models","author":"Daniel Chinonso John","date":"October 5, 2026","format":false,"excerpt":"Developers are turning retired FPGA hardware originally built for cryptocurrency mining into custom accelerators for Qwen3.5, including a Qwen3.5-9B implementation running on an SQRL Forest Kitten 33 board that sells for roughly $270 to $350 on the secondary market. The Forest Kitten 33, or FK33, uses a Xilinx Virtex UltraScale+\u2026","rel":"","context":"In &quot;Tech Updates&quot;","block_context":{"text":"Tech Updates","link":"https:\/\/areeblog.com\/category\/tech-updates\/"},"img":{"alt_text":"Developers Repurpose Crypto-Mining FPGAs to Run Qwen3.5 AI Models","src":"https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/10\/images-35.jpeg?resize=350%2C200&ssl=1","width":350,"height":200,"srcset":"https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/10\/images-35.jpeg?resize=350%2C200&ssl=1 1x, https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/10\/images-35.jpeg?resize=525%2C300&ssl=1 1.5x, https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/10\/images-35.jpeg?resize=700%2C400&ssl=1 2x"},"classes":[]},{"id":6575,"url":"https:\/\/areeblog.com\/london-ai-infrastructure-startup-callosum-raises-100-million-seed-round-led-by-atomico\/","url_meta":{"origin":7077,"position":1},"title":"London AI Infrastructure Startup Callosum Raises $100 Million Seed Round Led by Atomico","author":"Daniel Chinonso John","date":"August 21, 2026","format":false,"excerpt":"London-based artificial intelligence infrastructure startup Callosum has raised $100 million (\u20ac85.4 million) in a seed funding round led by Atomico, as the company develops software designed to route AI workloads across different models and computing hardware. The round also includes Plural, DCVC and the UK's Sovereign AI Fund. The financing\u2026","rel":"","context":"In &quot;Tech Updates&quot;","block_context":{"text":"Tech Updates","link":"https:\/\/areeblog.com\/category\/tech-updates\/"},"img":{"alt_text":"London AI Infrastructure Startup Callosum Raises $100 Million Seed Round Led by Atomico","src":"https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/08\/images-35.jpeg?resize=350%2C200&ssl=1","width":350,"height":200,"srcset":"https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/08\/images-35.jpeg?resize=350%2C200&ssl=1 1x, https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/08\/images-35.jpeg?resize=525%2C300&ssl=1 1.5x, https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/08\/images-35.jpeg?resize=700%2C400&ssl=1 2x"},"classes":[]},{"id":6281,"url":"https:\/\/areeblog.com\/where-ai-investment-is-moving-in-2026\/","url_meta":{"origin":7077,"position":2},"title":"Where AI Investment Is Moving in 2026","author":"Daniel Chinonso John","date":"July 10, 2026","format":false,"excerpt":"Artificial intelligence investment has entered a new phase in 2026. After years of excitement surrounding large language models and consumer-facing AI tools, investors are becoming more selective. The question is no longer who has the flashiest AI model. Instead, capital is flowing toward businesses that solve real problems, improve productivity,\u2026","rel":"","context":"In &quot;Artificial Intelligence&quot;","block_context":{"text":"Artificial Intelligence","link":"https:\/\/areeblog.com\/category\/artificial-intelligence\/"},"img":{"alt_text":"Where AI Investment Is Moving in 2026","src":"https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/07\/IMG-20260710-WA0014.jpg?resize=350%2C200&ssl=1","width":350,"height":200,"srcset":"https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/07\/IMG-20260710-WA0014.jpg?resize=350%2C200&ssl=1 1x, https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/07\/IMG-20260710-WA0014.jpg?resize=525%2C300&ssl=1 1.5x, https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/07\/IMG-20260710-WA0014.jpg?resize=700%2C400&ssl=1 2x, https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/07\/IMG-20260710-WA0014.jpg?resize=1050%2C600&ssl=1 3x"},"classes":[]},{"id":6771,"url":"https:\/\/areeblog.com\/trustkernel-ships-plugclaw-a-usb-c-ai-agent-computer\/","url_meta":{"origin":7077,"position":3},"title":"TrustKernel Ships PlugClaw, a USB-C AI Agent Computer","author":"Daniel Chinonso John","date":"September 5, 2026","format":false,"excerpt":"TrustKernel has begun worldwide shipments of PlugClaw, a small USB-C computer designed to give AI agents a separate computing environment from a user's phone, tablet or PC. The device was announced on August 12, 2026, and moved from preorder to worldwide general availability on September 4, according to TrustKernel. The\u2026","rel":"","context":"In &quot;Tech Updates&quot;","block_context":{"text":"Tech Updates","link":"https:\/\/areeblog.com\/category\/tech-updates\/"},"img":{"alt_text":"TrustKernel Ships PlugClaw, a USB-C AI Agent Computer","src":"https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/09\/images-53.jpeg?resize=350%2C200&ssl=1","width":350,"height":200,"srcset":"https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/09\/images-53.jpeg?resize=350%2C200&ssl=1 1x, https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/09\/images-53.jpeg?resize=525%2C300&ssl=1 1.5x, https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/09\/images-53.jpeg?resize=700%2C400&ssl=1 2x"},"classes":[]},{"id":5506,"url":"https:\/\/areeblog.com\/embodied-ai-how-machines-learn-by-doing\/","url_meta":{"origin":7077,"position":4},"title":"Embodied AI: How Machines Learn by Doing","author":"Samuel Ogori","date":"October 7, 2025","format":false,"excerpt":"Embodied AI means giving machines a body and the senses that come with it. Instead of only learning from text or images, these systems see, touch, move, and learn from what happens when they act. Think of a robot that learns to pick fruit by trying, feeling, and adjusting its\u2026","rel":"","context":"In &quot;Artificial Intelligence&quot;","block_context":{"text":"Artificial Intelligence","link":"https:\/\/areeblog.com\/category\/artificial-intelligence\/"},"img":{"alt_text":"Embodied AI: How Machines Learn by Doing","src":"https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2025\/10\/Embodied-AI.jpg?resize=350%2C200&ssl=1","width":350,"height":200,"srcset":"https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2025\/10\/Embodied-AI.jpg?resize=350%2C200&ssl=1 1x, https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2025\/10\/Embodied-AI.jpg?resize=525%2C300&ssl=1 1.5x, https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2025\/10\/Embodied-AI.jpg?resize=700%2C400&ssl=1 2x"},"classes":[]},{"id":6302,"url":"https:\/\/areeblog.com\/how-hallucinated-packages-become-an-attack-vector\/","url_meta":{"origin":7077,"position":5},"title":"How Hallucinated Packages Become an Attack Vector","author":"Daniel Chinonso John","date":"July 11, 2026","format":false,"excerpt":"Trust is becoming one of the most valuable\u2014and most dangerous\u2014currencies in software development. Every time an AI coding assistant suggests a library, most developers assume it exists. That assumption is increasingly being weaponized, not by breaking into software repositories, but by waiting for AI to imagine a package that has\u2026","rel":"","context":"In &quot;Cybersecurity&quot;","block_context":{"text":"Cybersecurity","link":"https:\/\/areeblog.com\/category\/cybersecurity\/"},"img":{"alt_text":"How Hallucinated Packages Become an Attack Vector","src":"https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/07\/IMG-20260712-WA0002.jpg?resize=350%2C200&ssl=1","width":350,"height":200,"srcset":"https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/07\/IMG-20260712-WA0002.jpg?resize=350%2C200&ssl=1 1x, https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/07\/IMG-20260712-WA0002.jpg?resize=525%2C300&ssl=1 1.5x, https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/07\/IMG-20260712-WA0002.jpg?resize=700%2C400&ssl=1 2x, https:\/\/i0.wp.com\/areeblog.com\/wp-content\/uploads\/2026\/07\/IMG-20260712-WA0002.jpg?resize=1050%2C600&ssl=1 3x"},"classes":[]}],"jetpack_featured_media_url":"https:\/\/areeblog.com\/wp-content\/uploads\/2026\/10\/ai-platform-blog-chat-labs-featured-1.png","_links":{"self":[{"href":"https:\/\/areeblog.com\/wp-json\/wp\/v2\/posts\/7077","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/areeblog.com\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/areeblog.com\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/areeblog.com\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/areeblog.com\/wp-json\/wp\/v2\/comments?post=7077"}],"version-history":[{"count":2,"href":"https:\/\/areeblog.com\/wp-json\/wp\/v2\/posts\/7077\/revisions"}],"predecessor-version":[{"id":7081,"href":"https:\/\/areeblog.com\/wp-json\/wp\/v2\/posts\/7077\/revisions\/7081"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/areeblog.com\/wp-json\/wp\/v2\/media\/7080"}],"wp:attachment":[{"href":"https:\/\/areeblog.com\/wp-json\/wp\/v2\/media?parent=7077"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/areeblog.com\/wp-json\/wp\/v2\/categories?post=7077"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/areeblog.com\/wp-json\/wp\/v2\/tags?post=7077"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}