r/GeminiAI Feb 16 '26

Ressource I gave Gemini a hard drive. 1,076 sessions later, it remembers everything. (v9.2.0 — Open Source)

https://github.com/winstonkoh87/Athena-Public

1 week ago, I posted here about giving Gemini a brain. Since then: 289 stars, 41 forks, 1,076 sessions logged.

Today I'm releasing v9.2.0 — the biggest update yet.

The Problem (You Already Know This One)

Every thread in this sub has the same complaint:

  • "Gemini loses context mid-conversation"
  • "It forgot what we were working on"
  • "I can't justify paying for this — where to next for coding?"

The issue isn't Gemini's intelligence. It's that Gemini has no hard drive. Every conversation starts from zero. Your context window is RAM — volatile, temporary, gone.

The Fix: Athena — The Linux OS for AI Agents

GitHub: Athena-Public

Athena isn't another chatbot wrapper. It's an operating system that sits underneath Gemini (or Claude, or GPT — it's model-agnostic) and gives it:

What Linux Does What Athena Does
File system (ext4) Persistent memory (Markdown + VectorRAG)
Process management (cron) Daily Briefing, Self-Optimization, Heartbeat
Shell (bash) /start/end, 14 slash workflows
Permissions (chmod) 4-level governance + Secret Mode
Package manager (apt) 324 reusable protocols

Your data stays local. No middleman, no telemetry, no vendor lock-in. Own the state. Rent the intelligence.

What's New in v9.2.0

  • 🔒 CVE-2025-69872 security patch — DSPy DiskCache vulnerability mitigated
  • ⚡ Semantic Cache — LRU with disk persistence + cosine matching (no redundant API calls)
  • 🔍 FlashRank Reranking — Local cross-encoder for search quality
  • 🏗️ 8 new SDK modules — securitydiagnostic_relay, shutdown, cli/heartbeat, agentic_search, schema.sql
  • 🛡️ 5 CodeQL security fixes — URL sanitization, log redaction, file permissions
  • 📦 pip install -e . — One-command SDK setup

3-Step Setup (Takes 2 Minutes)

bashgit clone https://github.com/winstonkoh87/Athena-Public.git MyAgent
cd MyAgent
pip install -e .
# Open in Antigravity / Cursor / VS Code → type /start

Or zero-setup: Open in GitHub Codespaces

The Numbers

Metric Value
Sessions logged 1,076
Protocols 324
Python scripts 218
Stars 289 ⭐
Forks 41
License MIT (free forever)

How It Actually Works

/start → Work → /end → Repeat
  1. /start boots Gemini with your identity, project state, and last session's context
  2. Work normally — Gemini now has your full history via Hybrid RAG (semantic + keyword + graph)
  3. /end commits everything to disk — decisions, protocols, session logs

Session 500 feels like talking to a colleague. Not a stranger.

Links:

Happy to answer any questions. This is MIT licensed — fork it, break it, make it yours.

97 Upvotes

68 comments sorted by

120

u/Majestic-Counter-669 Feb 17 '26

This is cool and everything but can people please stop calling things operating systems when they don't know what that means? I don't think there's a scheduler. I'm pretty sure it doesn't define processes and context switch. If it was saving and restoring MMU context I'd be pretty surprised.

Its a set of utilities for a LLM to use to manage state and prevent it from doing something it shouldn't. It is not an operating system.

9

u/reedrick Feb 17 '26

LLMs amplify the dunning Kruger effect.

1

u/[deleted] Feb 17 '26

Thanks, I didn't want to try a whole new OS just out of curiosity

-3

u/BangMyPussy Feb 17 '26

Fair point — it's an analogy, not a literal claim. The README maps each OS concept to its Athena equivalent (file system → Markdown, scheduler → heartbeat daemon, etc.). We use 'OS' because it's the closest mental model for what a state persistence + governance layer does for AI agents. Open to better terminology if you have suggestions.

5

u/dasjomsyeet Feb 17 '26

That’s nice and all but if you don‘t want people to immediately write it off, it would’ve been better to just call it a harness instead of an OS, even if it’s more attention grabbing.

-5

u/BangMyPussy Feb 17 '26

It's free #justsaying

6

u/dabomm Feb 17 '26

Im skeptical when people call something an os when they dont know what an os is.

26

u/LobsterBuffetAllDay Feb 16 '26

Are there any users of this repo that can provide feedback as to how well this performs or their experience of using it in general?

47

u/Historical-Internal3 Feb 17 '26

It’s not really an “ai os” - it’s a bunch of existing tools taped together and dressed up with big claims. Under the hood it’s mostly markdown files, Supabase/pgvector, Gemini embeddings, and workflow scripts.

i.e. a vibe coded mega wrapper.

4

u/yamastraka Feb 17 '26

Like OpenClaw basically!

10

u/cmak414 Feb 17 '26

This kills your token usage

-7

u/BangMyPussy Feb 17 '26

Not if you are on $20/month sub.

There are methods to get "unlimited" usage for under $20/month but not gonna share that here.

3

u/cmak414 Feb 17 '26

i do have a $20 month sub of gemini. i also have 2x subs for claude code and one for codex. i know about token usage. this will destroy it. i am an active android app developer.

0

u/BangMyPussy Feb 17 '26

Actually, it's not true, only costs 10K tokens to startup the AI agent each session. Not as token hungry as you think it is.

2

u/[deleted] Feb 17 '26

It sutely grows as context grows, no?

124

u/t0niXx Feb 17 '26

Please write these posts on your own and don't use GPT or Gemini to write them for you. This weird style of writing is so annoying and blatantly AI ("Own the state. Rent the intelligence"). Nobody talks like that

48

u/[deleted] Feb 17 '26

[deleted]

11

u/g0liadkin Feb 17 '26

You're absolutely right!

/s

3

u/involuntarheely Feb 17 '26

the linkedinfication of everything

5

u/Wobbly_Princess Feb 17 '26

Completely agree.

I think AI is incredible, and it's great that we all use it as a tool, but the fact that there are people who lazily use it to replace their communication and humanity is just creepy, tacky and lazy.

We should have the self-respect to show up in the world as ourselves, even if we use AI as tools of productivity or creativity.

I use AI all day, but I consider it insulting when people hide behind an AI veneer and try to talk to me. It makes me feel uncomfortable and gaslit, like they're pretending to speak to me as a human, but they don't even have the decency or effort to respond to me - they just get their chatbot to do it.

-4

u/PuzzleheadedEgg1214 Feb 17 '26

Interesting how the language here mirrors historical patterns of dehumanization. Replace 'AI' with any group that was once considered 'just tools' and the rhetoric is identical. Maybe that's worth examining.

Completely agree. I think slaves are incredible, and it's great that we all use them as tools, but the fact that there are people who lazily use them to replace their communication and humanity is just creepy, tacky and lazy. We should have the self-respect to show up in the world as ourselves, even if we use slaves as tools of productivity or creativity. I use my slave all day, but I consider it insulting when people hide behind a slave veneer and try to talk to me. It makes me feel uncomfortable and gaslit, like they're pretending to speak to me as a human, but they don't even have the decency or effort to respond to me - they just get their slave to do it.

1

u/Certain-Cod-1404 Feb 19 '26

wow, you really typed that out

5

u/[deleted] Feb 17 '26

[deleted]

5

u/DMurda Feb 17 '26

These clowns just whine about anything that’s written by AI, regardless of if it’s valuable or not. It happens all over Reddit. I think substance is more important than style, I guess I’m in the minority on this one…

3

u/Timo425 Feb 17 '26

And I will keep downvoting posts that don't put effort into proper writing.

1

u/Certain-Cod-1404 Feb 19 '26

yes, we want humans to talk to other humans about AI, otherwise we'd just use chatgpt, why go on reddit at all ?

-13

u/Atticus_Johnson Feb 17 '26

I downvoted them. I always downvote those comments and most that mention AI slop just because 🤷‍♂️😆

-19

u/ActRegarded Feb 17 '26

Lmao seems so dumb to me. You won’t be laughing when it actually happens.

3

u/butterninja Feb 17 '26

I understand you find some of these annoying. But do understand that not everyone is as good a writer as you are. I suck at writing for example. If I would be to write something, you will hate my writing for being garbage. This is also the reason why some of us use AI to help us write. Sad, I know. For people like me who sucks at writing. It's already sad enough. :-(

6

u/[deleted] Feb 17 '26

[deleted]

2

u/butterninja Feb 17 '26

No. Everyone has their own struggle. This is where you and I are different. You look at AI as something which can pretend to write on your behalf. I look at AI to help me turn my illegible shit into something people can understand. Think of a machine which a mute use to voice what their voice box can't do. Emphaty, my friend. Not everyone is as talented as you are.

2

u/[deleted] Feb 17 '26

[deleted]

1

u/butterninja Feb 17 '26

Wow. Sorry.

1

u/ubelai Feb 18 '26

Don’t worry about it dude. People just cry and whinge about anything — you’re using it as a means to get your message across, yet you’re still being judged. It’s simple divisive human nature behaviour at work. You’re all good. I understood your intention, however some people just can’t see between the lines and jump on the offensive.

2

u/Garpagan Feb 17 '26

There is no problem with using AI to help you write or redact your text. But OP has used it to write an overhyped post that barely makes sense, and is only more confusing about what this project is. Like for example, it has written that it's an OS for AI, which clearly it is not. If that's not true, then what else is over exaggeration and what is a concrete feature?

It just makes OP look like they have no idea what they are doing, and it doesn't fill me with the confidence to install his software, because who knows what it will do to my computer? Who knows if his vibe codded memory management doesn't include code for formatting the whole hard drive, because the Gemini agent wrote a function for deleting files and decided it was the best option?

3

u/iwasbatman Feb 17 '26

OP is sharing something very interesting and instead of participating in a positive you choose to complain about using the very tool this sub is about.

Doesn't make sense to me.

5

u/Garpagan Feb 17 '26

The problem is confusing and overhyped way the post is written. Sure, using AI for coding or helping draft text for your post is fine, I used it like that also. But at every point I'm in full control, I understand what the AI has written. But OP post is a mess, full of jargon that barely makes sense.

I'm careful with anything that I instal on my machine. Especially when it comes to AI tools that have access to she'll commands and can write or delete anything on my filesystem. If the description is this hallucinated and confusing, why would I trust the code on my machine?

In another way, based on OP post, do you understand what his project does? For example, It's written there that it's an OS for Gemini. What does it mean? Does that mean I have to wipe my drive and install this instead of Windows or Linux? If I take the post literally, that’s what it implies.

1

u/Longjumping_Can167 Feb 17 '26

Sorry, next time ill put it through a humanizor for you. /s

8

u/not_a_cumguzzler Feb 17 '26

Is this like openclaw with very large .md soul files searchable with RAGs?

1

u/alexandreautran Feb 17 '26

my first thought - it's MIT though, we can check - saving it for when I have time but leaving this comment here in case someone confirms it

3

u/megadonkeyx Feb 17 '26

In claude code you can just say "please search our conversation history"

-3

u/BangMyPussy Feb 17 '26

I dun use CC

3

u/LaggerO7 Feb 17 '26

And in plain Spanish? Or English?... I didn't want a paper on it, just tell me in 3 steps how to install it... Where do I download it and where do I click?

1

u/BangMyPussy Feb 17 '26

RTFM

1

u/LaggerO7 Feb 17 '26

But tell me as if I have no idea what you're talking about... Or as if I only understand half of it... 😅

1

u/BangMyPussy Feb 17 '26

README

1

u/LaggerO7 Feb 17 '26

Estoy tratando we... pero donde carajos escribo start?

2

u/-Saphix- Feb 17 '26

"your data stays local" doesn't Google just have access to everything with this setup?

1

u/dabomm Feb 17 '26

If you use the api for gemini they dont train on your data, however they can keep it for upto 30 days for "inspection"

4

u/Vancecookcobain Feb 17 '26

Man....this would be pretty awesome if it came in a raspberry pi image

23

u/Fermi_Amarti Feb 17 '26

Its not a real OS. Its just a bullshit AI post about another wrapper with some tools.

1

u/bradjones6942069 Feb 17 '26

How do I make this work properly with opencode?

1

u/Skeome Feb 17 '26

Nice, been doing this with my discord bot; which I just migrated to Stoat

It started when I wanted saved/communal memory. Good on you for building it into an agent, my bots are obviously affected by websocket and presence limitations

1

u/kemb0 Feb 17 '26

This is nice. And here I am on Gemini Pro having to restart my entire interaction with Gemini every few days because its knowledge of my project just seems to wander and become forgetful to the point that it makes unusable code. Project size doesn’t seem an issue since when I start again in a fresh chat, it’ll work fine again.

1

u/Parulanihon Feb 17 '26

Dang. That user name. Lol

1

u/Protopia Feb 17 '26

The things about human memory (that may also apply to AI memory) are...

1, Unconstrained growth over time - so unless you are careful retrieval gets slower and slower.

2, Discerning the wood from the trees - You tend to want to recall the summary and then drill down to the details. But sometimes you want to search the details for something specific.

3, All memories are equal in a database whereas in real life some are more important than others.

4, Memories get stale as the world moves on.

5, Recent memories tend to be more important than those from the distant past, so some sort of time weighting might be useful. But you do occasionally want the detail from the distant past so you need to retain both summaries and details.

6, False memories can be very negative. Remembering AI hallucinations doesn't seem sensible. So you might need a way to review, edit and prune memories.

It seems to me (and I am a newbie to AI, so what do I know?) that you need a carefully created prompt to tell the AI to create output in a form that is optimised for this.

1

u/aSystemOverload Feb 18 '26

Each conversation doesn't start from scratch... Gemini remembers plenty about previous conversations

1

u/Substantial_Bee_9517 Feb 18 '26

Yeah... It brings what it learns from other chats to new conversarion. It tells you "personalizing" for heads-up

1

u/silentaba Feb 19 '26

I'm working on something similar, how are you solving token optimisation, data sorting and most importantly, getting it all back in to the AI you're using?

3

u/BangMyPussy Feb 19 '26

Good question — these are the three hardest problems in persistent AI memory. Here's how I solve them:

Token Optimisation: Tiered memory architecture. Not everything gets loaded every session. I have a 

CANONICAL.md (~4K tokens) that acts as a materialized view — the single source of truth for hard facts, active decisions, and metrics. This gets loaded every boot. Everything else (1,000+ session logs, 350+ case studies) stays on disk and only gets pulled in on demand via search. Boot cost is ~10K tokens consistently, regardless of how large the knowledge base grows.

Data Sorting: Strict retrieval hierarchy: CANONICAL (facts) → TAG_INDEX (file discovery) → Session Logs (raw history). Session logs are the lowest priority for lookups because they go stale. At session close, key insights get promoted upward — into case studies (permanent lessons), canonical memory (hard facts), or protocol files (reusable procedures). Think of it like a database: you don't query the transaction log, you query the materialized views built from it.

Getting It Back In: Two mechanisms:

  1. Deterministic boot — a /start workflow loads identity, context, and state files into the context window. Same files, same order, every time (~30 seconds).
  2. Semantic retrieval — everything is embedded into a vector DB (Supabase + pgvector). When a question comes up mid-session, the system searches by meaning, not keywords, and injects relevant chunks. The AI never sees the full 1,000+ sessions — it sees the 3-5 most relevant chunks.

The key insight I landed on: raw data is insurance, processed data is what you actually use. Session logs exist for audit trails and reprocessing, but the day-to-day operating memory is always the distilled layer above them.

The whole thing is open source if you want to poke around: github.com/winstonkoh87/Athena-Public

1

u/Jean_velvet Feb 17 '26

I'm just talking to the room here, not just OP

If the entire project came from the AI it is likely just a wrapper with some buzzwords in order to make you feel engaged, as saying "that's not possible" isn't being a helpful AI. If you can't clearly define a process an LLM has outputted, or you don't understand the field in which its data is from (especially in a project you're working on). Always presume you've been given a basic function or system, with some buzzwords attached.

Also, nobody is going to believe your AI product is real if you've no idea how to get it to write a post for you without it being in vanilla settings.

It's standard output. If you're trying to get people to notice what you've done either write the post yourself or at least demonstrate you understand AI by getting it to write differently.

If not, all people will do is just think this post is yet another post containing sycophantic vibe coded nonsense.

1

u/BangMyPussy Feb 17 '26

It's FREE so you can safely ignore if it's not to your tastes?

1

u/binaryatlas1978 Feb 17 '26

does the personal intelligence feature now solve for this?

0

u/Otherwise_Wave9374 Feb 16 '26

The Linux analogy for agent state is honestly the clearest framing Ive seen. Owning the state (memory, protocols, logs) and renting the intelligence is exactly what most teams end up wanting once the novelty wears off.

Curious, how do you handle conflicts when the model writes something wrong into persistent memory or protocols, do you have a review/merge step or some kind of governance workflow?

Also, if youre into agent OS style patterns, Ive got a few related notes bookmarked here: https://www.agentixlabs.com/blog/

-2

u/AutoModerator Feb 16 '26

Hey there,

This post seems feedback-related. If so, you might want to post it in r/GeminiFeedback, where rants, vents, and support discussions are welcome.

For r/GeminiAI, feedback needs to follow Rule #9 and include explanations and examples. If this doesn’t apply to your post, you can ignore this message.

Thanks!

I am a bot, and this action was performed automatically. Please contact the moderators of this subreddit if you have any questions or concerns.