{"id":306972,"date":"2026-07-01T17:10:37","date_gmt":"2026-07-01T17:10:37","guid":{"rendered":"https:\/\/aiassetman.com\/?p=306972"},"modified":"2026-07-01T17:10:37","modified_gmt":"2026-07-01T17:10:37","slug":"how-much-data-can-an-llm-hold-only-for-techies","status":"publish","type":"post","link":"https:\/\/aiassetman.com\/how-much-data-can-an-llm-hold-only-for-techies\/","title":{"rendered":"How Much Data Can an LLM hold (only for techies)"},"content":{"rendered":"<p>I queried Google.com with the title question and got this from it&#8217;s AI.<\/p>\n<p>======================================================<\/p>\n<div class=\"n6owBd awi2gc\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCA4QAA\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px 0px 16px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><mark class=\"HxTRcb\" data-sfc-root=\"ep\" data-wiz-uids=\"r1oD2e_k\" data-sfc-cb=\"\" data-ved=\"2ahUKEwjTgo_I-rGVAxWa4skDHdghCtoQuJAPegoIAggACAEIDhAB\" data-complete=\"true\" data-sfc-inited=\"2\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 500; margin: 0px; text-decoration: none; border-bottom: 0px rgb(0, 29, 53);\"><span data-subtree=\"aimfl,mfl\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 500; margin: 0px; text-decoration: none; border-bottom: 0px rgb(0, 29, 53);\">LLMs do not &#8220;store&#8221; information in a traditional database<\/span><!--TgQPHd||[]--><\/mark> [<a href=\"https:\/\/iapp.org\/news\/a\/do-llms-store-personal-data-this-is-asking-the-wrong-question\">1<\/a>, <a href=\"https:\/\/laweconcenter.org\/resources\/llms-are-not-databases-memorization-disclosure-and-the-limits-of-privacy-law\/\">2<\/a>]. Instead, their knowledge is either permanently baked into their internal parameters (the model&#8217;s &#8220;brain&#8221;) during training, or loaded temporarily into their <strong class=\"Yjhzub\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 700; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">context window<!--TgQPHd||[]--><\/strong> (their &#8220;working memory&#8221;) during a prompt. [<a href=\"https:\/\/iapp.org\/news\/a\/do-llms-store-personal-data-this-is-asking-the-wrong-question\">1<\/a>, <a href=\"https:\/\/laweconcenter.org\/resources\/llms-are-not-databases-memorization-disclosure-and-the-limits-of-privacy-law\/\">2<\/a>, <a href=\"https:\/\/help.flintk12.com\/en\/articles\/9025323-what-knowledge-does-ai-have-access-to\">3<\/a>, <a href=\"https:\/\/medium.com\/data-science-collective\/inside-the-context-window-how-llms-actually-remember-054d763e47b8\">4<\/a>]<!--TgQPHd||[]--><\/div>\n<div class=\"Fsg96\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-sfc-inited=\"2\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><!--TgQPHd||[]--><\/div>\n<div class=\"otQkpb\" role=\"heading\" aria-level=\"3\" data-animation-nesting=\"\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 20px; font-weight: 600; margin: 24px 0px 12px; text-decoration: none; border-bottom: 0px rgb(0, 29, 53);\"><strong>1. The Context Window (Working Memory)<\/strong><!--TgQPHd||[]--><\/div>\n<div class=\"\" data-bfc=\"\" data-ved=\"2ahUKEwjTgo_I-rGVAxWa4skDHdghCtoQi4wTegoIAggACAEIEBAA\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">\n<div class=\"n6owBd awi2gc\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCBAQAQ\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 12px 0px 16px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">This is the amount of text an AI can process in a single prompt or conversation. Because 1,000 tokens equal roughly 750 English words, modern AI capacities are immense: [<a href=\"https:\/\/www.reddit.com\/r\/LocalLLaMA\/comments\/144ch8y\/please_help_me_understand_the_limitations_of\/\">1<\/a>, <a href=\"https:\/\/www.ibm.com\/think\/topics\/context-window\">2<\/a>, <a href=\"https:\/\/atlan.com\/know\/llm-context-window-limitations\/\">3<\/a>, <a href=\"https:\/\/www.siliconflow.com\/articles\/en\/top-LLMs-for-long-context-windows\">4<\/a>]<!--TgQPHd||[]--><\/div>\n<\/div>\n<div class=\"\" data-bfc=\"\" data-ved=\"2ahUKEwjTgo_I-rGVAxWa4skDHdghCtoQi4wTegoIAggACAEIHhAA\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">\n<ul class=\"KsbFXc U6u95\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 12px 0px 16px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">\n<li class=\"Z1qcYe\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCB4QAQ\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px 0px 12px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><span class=\"T286Pc\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><strong class=\"Yjhzub\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 700; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">Standard Models (e.g., GPT-4o, Claude 3.5 Sonnet):<!--TgQPHd||[]--><\/strong> Typically hold 128k to 200k tokens, which is roughly <strong class=\"Yjhzub\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 700; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">90,000 to 150,000 words<!--TgQPHd||[]--><\/strong> (equal to 1 to 2 average-length books).<!--TgQPHd||[]--><\/span> [<a href=\"https:\/\/medium.com\/@aloy.banerjee30\/infinite-context-length-in-llms-the-next-big-advantage-in-ai-2550e9e6ce9b\">1<\/a>, <a href=\"https:\/\/pristren.com\/blog\/llm-context-window-comparison\/\">2<\/a>, <a href=\"https:\/\/local-ai-zone.github.io\/guides\/context-length-optimization-ultimate-guide-2025.html\">3<\/a>]<!--TgQPHd||[]--><\/li>\n<li class=\"Z1qcYe\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCB4QBw\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px 0px 12px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><span class=\"T286Pc\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><strong class=\"Yjhzub\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 700; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">Ultra-Long Context Models (e.g., Gemini 1.5\/3 Pro):<!--TgQPHd||[]--><\/strong> Process up to 1 to 2 million tokens, or roughly <strong class=\"Yjhzub\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 700; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">750,000 to 1.5 million words<!--TgQPHd||[]--><\/strong>. This allows the AI to &#8220;read&#8221; hundreds of files or entire code repositories in one go.<!--TgQPHd||[]--><\/span> [<a href=\"https:\/\/localaimaster.com\/models\/context-windows-coding-explained\">1<\/a>, <a href=\"https:\/\/gist.github.com\/LEX8888\/3f4183df6fef0d6e4783aae1bd986d17\">2<\/a>, <a href=\"https:\/\/www.siliconflow.com\/articles\/en\/top-LLMs-for-long-context-windows\">3<\/a>, <a href=\"https:\/\/local-ai-zone.github.io\/guides\/context-length-optimization-ultimate-guide-2025.html\">4<\/a>]<!--TgQPHd||[]--><\/li>\n<p><!--TgQPHd||[]--><\/ul>\n<\/div>\n<div class=\"Fsg96\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-sfc-inited=\"2\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><!--TgQPHd||[]--><\/div>\n<div class=\"\" data-bfc=\"\" data-ved=\"2ahUKEwjTgo_I-rGVAxWa4skDHdghCtoQi4wTegoIAggACAEIHxAA\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">\n<div class=\"otQkpb\" role=\"heading\" aria-level=\"3\" data-animation-nesting=\"\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 20px; font-weight: 600; margin: 24px 0px 12px; text-decoration: none; border-bottom: 0px rgb(0, 29, 53);\"><strong>2. Internal Memory (Trained Knowledge)<\/strong><!--TgQPHd||[]--><\/div>\n<\/div>\n<div class=\"\" data-bfc=\"\" data-ved=\"2ahUKEwjTgo_I-rGVAxWa4skDHdghCtoQi4wTegoIAggACAEIIBAA\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">\n<div class=\"n6owBd awi2gc\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCCAQAQ\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 12px 0px 16px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">Beyond temporary prompts, an LLM&#8217;s vast general knowledge is encoded mathematically into billions of <strong class=\"Yjhzub\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 700; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">parameters<!--TgQPHd||[]--><\/strong> (the connections between artificial neurons). [<a href=\"https:\/\/www.facebook.com\/groups\/DeepNetGroup\/posts\/1984705291922343\/\">1<\/a>, <a href=\"https:\/\/read.dukeupress.edu\/critical-ai\/article\/doi\/10.1215\/2834703X-11556011\/400182\/Eight-Things-to-Know-about-Large-Language-Models\">2<\/a>]<!--TgQPHd||[]--><\/div>\n<\/div>\n<div class=\"\" data-bfc=\"\" data-ved=\"2ahUKEwjTgo_I-rGVAxWa4skDHdghCtoQi4wTegoIAggACAEIIxAA\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">\n<ul class=\"KsbFXc U6u95\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 12px 0px 16px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">\n<li class=\"Z1qcYe\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCCMQAQ\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px 0px 12px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><span class=\"T286Pc\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">A mid-sized model has tens of billions of parameters.<!--TgQPHd||[]--><\/span><!--TgQPHd||[]--><\/li>\n<li class=\"Z1qcYe\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCCMQAg\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px 0px 12px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><span class=\"T286Pc\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">This allows the AI to &#8220;memorize&#8221; global facts, language rules, and vast amounts of public information, but it means the model cannot selectively &#8220;forget&#8221; or update its knowledge without a retraining cycle.<!--TgQPHd||[]--><\/span> [<a href=\"https:\/\/help.flintk12.com\/en\/articles\/9025323-what-knowledge-does-ai-have-access-to\">1<\/a>, <a href=\"https:\/\/www.informationweek.com\/machine-learning-ai\/is-a-small-language-model-better-than-an-llm-for-you-\">2<\/a>, <a href=\"https:\/\/apxml.com\/courses\/intro-large-language-models\/chapter-4-different-llm-types\/model-size-and-capabilities\">3<\/a>, <a href=\"https:\/\/www.reddit.com\/r\/ArtificialInteligence\/comments\/1lriwdn\/what_my_views_are_for_people_worried_about\/\">4<\/a>, <a href=\"https:\/\/www.mdpi.com\/1999-5903\/17\/4\/151\">5<\/a>]<!--TgQPHd||[]--><\/li>\n<p><!--TgQPHd||[]--><\/ul>\n<\/div>\n<div class=\"Fsg96\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-sfc-inited=\"2\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><!--TgQPHd||[]--><\/div>\n<div class=\"\" data-bfc=\"\" data-ved=\"2ahUKEwjTgo_I-rGVAxWa4skDHdghCtoQi4wTegoIAggACAEIJhAA\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">\n<div class=\"otQkpb\" role=\"heading\" aria-level=\"3\" data-animation-nesting=\"\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 20px; font-weight: 600; margin: 24px 0px 12px; text-decoration: none; border-bottom: 0px rgb(0, 29, 53);\"><strong>3. External Memory (Databases &amp; RAG)<\/strong><!--TgQPHd||[]--><\/div>\n<\/div>\n<div class=\"\" data-bfc=\"\" data-ved=\"2ahUKEwjTgo_I-rGVAxWa4skDHdghCtoQi4wTegoIAggACAEIJxAA\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">\n<div class=\"n6owBd awi2gc\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCCcQAQ\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 12px 0px 16px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">Because LLM context windows fill up and are volatile, developers and businesses use <strong class=\"Yjhzub\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 700; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">Retrieval-Augmented Generation (RAG)<!--TgQPHd||[]--><\/strong>. RAG connects an LLM to an external vector database. [<a href=\"https:\/\/medium.com\/data-science-collective\/inside-the-context-window-how-llms-actually-remember-054d763e47b8\">1<\/a>, <a href=\"https:\/\/memgraph.com\/blog\/llm-limitations-query-enterprise-data\">2<\/a>, <a href=\"https:\/\/airbyte.com\/agentic-data\/large-context-window\">3<\/a>, <a href=\"https:\/\/www.chitika.com\/how-to-use-rag-in-llm\/\">4<\/a>]<!--TgQPHd||[]--><\/div>\n<\/div>\n<div class=\"\" data-bfc=\"\" data-ved=\"2ahUKEwjTgo_I-rGVAxWa4skDHdghCtoQi4wTegoIAggACAEIKRAA\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">\n<ul class=\"KsbFXc U6u95\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 12px 0px 16px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">\n<li class=\"Z1qcYe\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCCkQAQ\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px 0px 12px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><span class=\"T286Pc\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">Instead of forcing the AI to memorize everything, the AI acts as a search engine operator, querying external databases to find the exact documents it needs, and then pulling only those specific documents into its temporary context window to answer your question.<!--TgQPHd||[]--><\/span> [<a href=\"https:\/\/airbyte.com\/agentic-data\/large-context-window\">1<\/a>, <a href=\"https:\/\/medium.com\/@Micheal-Lanham\/rag-how-ai-gets-an-open-book-to-stop-making-things-up-2907dab801e7\">2<\/a>, <a href=\"https:\/\/medium.com\/data-science-collective\/inside-the-context-window-how-llms-actually-remember-054d763e47b8\">3<\/a>, <a href=\"https:\/\/www.sistrix.com\/ask-sistrix\/ai-basics\/ai-search-engines-how-to-ensure-visibility-in-the-era-of-ai-search\/\">4<\/a>]<!--TgQPHd||[]--><\/li>\n<p><!--TgQPHd||[]--><\/ul>\n<\/div>\n<div class=\"Fsg96\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-sfc-inited=\"2\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><!--TgQPHd||[]--><\/div>\n<div class=\"\" data-bfc=\"\" data-ved=\"2ahUKEwjTgo_I-rGVAxWa4skDHdghCtoQi4wTegoIAggACAEIKhAA\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">\n<div data-sfc-cp=\"\" data-sfc-root=\"c\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCCoQAQ\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">\n<div class=\"n6owBd awi2gc\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCCoQAg\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 12px 0px 16px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">If you are looking to build an AI system or want to manage a massive amount of your own data, tell me:<\/div>\n<ul class=\"KsbFXc U6u95\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 12px 0px 16px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">\n<li class=\"Z1qcYe\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCCoQAw\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px 0px 12px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><span class=\"T286Pc\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><strong class=\"Yjhzub\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 700; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">What type of information<\/strong> you need the AI to access (e.g., PDFs, a company website, codebases)<\/span><\/li>\n<li class=\"Z1qcYe\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCCoQBA\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px 0px 12px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><span class=\"T286Pc\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\"><strong class=\"Yjhzub\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 700; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">How much data<\/strong> you estimate you have (in gigabytes or total document count)<\/span><\/li>\n<\/ul>\n<div class=\"n6owBd awi2gc\" data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCCoQBQ\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 12px 0px 16px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">I can recommend the <strong class=\"Yjhzub\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 700; margin: 0px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">best tools or frameworks<\/strong> for setting up your AI storage!<\/p>\n<hr \/>\n<p>End of AI answer.\u00a0 I just reposted because it&#8217;s interesting to me.\u00a0 If you need more info, the blue numbers are links to further info.<\/p><\/div>\n<\/div>\n<div data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCCoQBQ\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 12px 0px 16px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">All the info above came from Google&#8217;s Gemini.<\/div>\n<\/div>\n<div data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCCoQBQ\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 12px 0px 16px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">\u00a0<\/div>\n<div data-sfc-cp=\"\" data-sfc-root=\"ep\" data-sfc-cb=\"\" data-hveid=\"CAIIAAgBCCoQBQ\" data-complete=\"true\" data-copy-service-computed-style=\"font-family: &quot;Google Sans&quot;, Roboto, Arial, sans-serif; font-size: 16px; font-weight: 400; margin: 12px 0px 16px; text-decoration: none; border-bottom: 0px rgb(10, 10, 10);\">Fred\u00a0<em>The Submarine Guy <\/em>Raley<\/div>\n","protected":false},"excerpt":{"rendered":"<p>I queried Google.com with the title question and got this from it&#8217;s AI. ====================================================== LLMs do not &#8220;store&#8221; information in a traditional database [1, 2]. Instead, their knowledge is either permanently baked into their internal parameters (the model&#8217;s &#8220;brain&#8221;) during training, or loaded temporarily into their context window (their &#8220;working memory&#8221;) during a prompt. [1, [&hellip;]<\/p>\n","protected":false},"author":2,"featured_media":0,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"order-bump-settings":[],"_wpfnl_thankyou_order_overview":"on","_wpfnl_thankyou_order_details":"on","_wpfnl_thankyou_billing_details":"on","_wpfnl_thankyou_shipping_details":"on","footnotes":""},"categories":[1],"tags":[],"class_list":["post-306972","post","type-post","status-publish","format-standard","hentry","category-affiliate-marketing"],"_links":{"self":[{"href":"https:\/\/aiassetman.com\/v\/wp\/v2\/posts\/306972","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/aiassetman.com\/v\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/aiassetman.com\/v\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/aiassetman.com\/v\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/aiassetman.com\/v\/wp\/v2\/comments?post=306972"}],"version-history":[{"count":1,"href":"https:\/\/aiassetman.com\/v\/wp\/v2\/posts\/306972\/revisions"}],"predecessor-version":[{"id":306973,"href":"https:\/\/aiassetman.com\/v\/wp\/v2\/posts\/306972\/revisions\/306973"}],"wp:attachment":[{"href":"https:\/\/aiassetman.com\/v\/wp\/v2\/media?parent=306972"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/aiassetman.com\/v\/wp\/v2\/categories?post=306972"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/aiassetman.com\/v\/wp\/v2\/tags?post=306972"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}