<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:dc="http://purl.org/dc/elements/1.1/">
  <channel>
    <title>DEV Community: Prathamesh Guram</title>
    <description>The latest articles on DEV Community by Prathamesh Guram (@prathameshguram2007).</description>
    <link>https://dev.to/prathameshguram2007</link>
    <image>
      <url>https://media2.dev.to/dynamic/image/width=90,height=90,fit=cover,gravity=auto,format=auto/https:%2F%2Fthepracticaldev.s3.amazonaws.com%2Fi%2F99mvlsfu5tfj9m7ku25d.png</url>
      <title>DEV Community: Prathamesh Guram</title>
      <link>https://dev.to/prathameshguram2007</link>
    </image>
    <atom:link rel="self" type="application/rss+xml" href="https://dev.to/feed/prathameshguram2007"/>
    <language>en</language>
    <item>
      <title>VoxBuddy</title>
      <dc:creator>Prathamesh Guram</dc:creator>
      <pubDate>Sun, 04 Oct 2026 09:26:59 +0000</pubDate>
      <link>https://dev.to/prathameshguram2007/voxbuddy-1hg6</link>
      <guid>https://dev.to/prathameshguram2007/voxbuddy-1hg6</guid>
      <description>&lt;h2&gt;
  
  
  What I Built
&lt;/h2&gt;

&lt;p&gt;I built my friend a voice-note assistant that keeps their audio on their laptop&lt;br&gt;
My friend has a habit I think a lot of us have: they record voice notes whenever something comes to mind, then almost never come back to them.&lt;br&gt;
assignments. People to call. Things to buy. Meetings. Random details they know they will forget.&lt;/p&gt;

&lt;p&gt;So for this Hacktoberfest weekend, I built &lt;strong&gt;VoxBuddy&lt;/strong&gt; — a small local AI app that turns a messy voice note into a clear summary, actionable tasks, priorities, and important details.&lt;/p&gt;
&lt;h2&gt;
  
  
  Demo
&lt;/h2&gt;

&lt;p&gt;  &lt;iframe src="https://www.youtube.com/embed/OW3o6PIhBaU" width="710" height="399"&gt;
  &lt;/iframe&gt;
&lt;/p&gt;

&lt;h2&gt;
  
  
  Code
&lt;/h2&gt;


&lt;div class="ltag-github-readme-tag"&gt;
  &lt;div class="readme-overview"&gt;
    &lt;h2&gt;
      &lt;img src="https://assets.dev.to/assets/github-logo-5a155e1f9a670af7944dd5e12375bc76ed542ea80224905ecaf878b9157cdefc.svg" alt="GitHub logo"&gt;
      &lt;a href="https://github.com/Prathamesh2007-o7" rel="noopener noreferrer"&gt;
        Prathamesh2007-o7
      &lt;/a&gt; / &lt;a href="https://github.com/Prathamesh2007-o7/voxbuddy" rel="noopener noreferrer"&gt;
        voxbuddy
      &lt;/a&gt;
    &lt;/h2&gt;
    &lt;h3&gt;
      
    &lt;/h3&gt;
  &lt;/div&gt;
  &lt;div class="ltag-github-body"&gt;
    
&lt;div id="readme" class="md"&gt;&lt;div class="markdown-heading"&gt;
&lt;h1 class="heading-element"&gt;VoxBuddy 🎙️&lt;/h1&gt;
&lt;/div&gt;
&lt;p&gt;&lt;strong&gt;Turn messy voice notes into things you can act on — privately, with local open AI.&lt;/strong&gt;&lt;/p&gt;
&lt;p&gt;VoxBuddy was built for a friend who records lots of voice notes but rarely revisits them. Instead of sending those recordings to a cloud AI API, VoxBuddy keeps the AI pipeline on the user's machine:&lt;/p&gt;
&lt;p&gt;&lt;strong&gt;Voice note → Whisper → transcript → Qwen3 → summary + tasks + important details&lt;/strong&gt;&lt;/p&gt;
&lt;div class="markdown-heading"&gt;
&lt;h2 class="heading-element"&gt;Why this project exists&lt;/h2&gt;
&lt;/div&gt;
&lt;p&gt;A voice note is easy to record and surprisingly easy to forget.&lt;/p&gt;
&lt;p&gt;VoxBuddy takes the few seconds your friend already spends speaking and turns them into something actionable:&lt;/p&gt;
&lt;ul&gt;
&lt;li&gt;A short summary&lt;/li&gt;
&lt;li&gt;Clear tasks&lt;/li&gt;
&lt;li&gt;Priority levels&lt;/li&gt;
&lt;li&gt;Deadlines when they were actually mentioned&lt;/li&gt;
&lt;li&gt;The original transcript for verification&lt;/li&gt;
&lt;/ul&gt;
&lt;p&gt;The goal is not to create another general-purpose chatbot. It is a tiny tool built around one real person's habit.&lt;/p&gt;
&lt;div class="markdown-heading"&gt;
&lt;h2 class="heading-element"&gt;Why open AI matters here&lt;/h2&gt;
&lt;/div&gt;
&lt;p&gt;The most important design choice is that…&lt;/p&gt;&lt;/div&gt;
  &lt;/div&gt;
  &lt;div class="gh-btn-container"&gt;&lt;a class="gh-btn" href="https://github.com/Prathamesh2007-o7/voxbuddy" rel="noopener noreferrer"&gt;View on GitHub&lt;/a&gt;&lt;/div&gt;
&lt;/div&gt;


&lt;h2&gt;
  
  
  How I Built It
&lt;/h2&gt;

&lt;p&gt;I built VoxBuddy as a local-first AI application using open-source AI models and local inference.&lt;/p&gt;

&lt;p&gt;The app starts with a voice note recorded or uploaded through the browser. The audio is sent to a &lt;strong&gt;Flask backend&lt;/strong&gt;, where &lt;strong&gt;Whisper&lt;/strong&gt; runs locally to convert the speech into text. This gives VoxBuddy an accurate transcript without sending the recording to a cloud speech API.&lt;/p&gt;

&lt;p&gt;The transcript is then passed to &lt;strong&gt;Qwen3 (4B)&lt;/strong&gt; running locally through &lt;strong&gt;Ollama&lt;/strong&gt;. I prompt the model to turn the unstructured voice note into structured information such as a summary, actionable tasks, priorities, and deadlines. The backend validates the model's response and extracts the JSON before sending the results back to the frontend.&lt;/p&gt;

&lt;p&gt;The main flow is:&lt;/p&gt;

&lt;p&gt;&lt;strong&gt;Voice note → Local Whisper → Transcript → Local Qwen3/Ollama → Structured tasks &amp;amp; summary → Web UI&lt;/strong&gt;&lt;/p&gt;

&lt;p&gt;I used &lt;strong&gt;HTML, CSS, and JavaScript&lt;/strong&gt; for the frontend and &lt;strong&gt;Python + Flask&lt;/strong&gt; for the backend. Everything is designed to run on the user's own laptop once the models and dependencies are installed.&lt;/p&gt;

&lt;p&gt;The open-source/local approach is an important part of the project rather than just a technical choice: voice notes can contain personal or sensitive information, so VoxBuddy can process them locally without requiring a third-party AI server. It also means the models can be swapped, modified, or upgraded without redesigning the entire application.&lt;/p&gt;

&lt;h2&gt;
  
  
  Why Does Open Innovation Matter?
&lt;/h2&gt;

&lt;ul&gt;
&lt;li&gt;
&lt;strong&gt;Privacy:&lt;/strong&gt; Voice notes can contain personal information, and VoxBuddy processes them locally instead of sending them to a cloud AI service.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Offline use:&lt;/strong&gt; After setup, the core AI processing can run without an internet connection.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;No per-request API costs:&lt;/strong&gt; Local inference avoids ongoing cloud API charges.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Model freedom:&lt;/strong&gt; I can swap or experiment with different open models instead of being locked into one provider.&lt;/li&gt;
&lt;li&gt;
&lt;strong&gt;Control:&lt;/strong&gt; I can change the prompts, processing pipeline, and application behavior to fit the user's needs.&lt;/li&gt;
&lt;/ul&gt;

&lt;h2&gt;
  
  
  Prize Categories
&lt;/h2&gt;

&lt;p&gt;&lt;strong&gt;Overall Winner&lt;/strong&gt;&lt;/p&gt;

</description>
      <category>devchallenge</category>
      <category>weekendchallenge</category>
      <category>hf26challenge</category>
    </item>
  </channel>
</rss>
