Back

Why your AI chatbot suddenly gets stupid mid-conversation

Discover why AI chatbots like ChatGPT and Claude suddenly lose the plot mid-conversation. Learn what context windows are and how to fix it, with ivee.

Date

Reading time

7

min

Amelia Miller

Co-founder and CEO

You are halfway through a really productive conversation with ChatGPT or Claude. The outputs are great. Then suddenly, out of nowhere, it starts contradicting itself, ignoring your instructions and giving you the most generic responses imaginable. Sound familiar? You have not done anything wrong. Your AI chatbot has simply run out of memory. This is what people call a context window, and once you understand how it works, you will never waste a conversation again. Here is everything you need to know about AI chatbot memory limits, plus the practical fixes that actually work.

What is an AI chatbot memory limit?

Every AI chatbot has a memory limit. It is not stored in a hard drive somewhere. It is measured in tokens, which are small chunks of text roughly equivalent to three quarters of a word. Every message you send and every response the AI generates uses up tokens. Once you hit the limit, the AI literally forgets the beginning of your conversation and starts working only from what it can still see.

This is what people call a context window. Think of it like a sliding window moving along your conversation. Everything inside the window is what the AI can see and use. Everything outside it is gone.

How big are the context windows for popular AI chatbots?

ChatGPT’s context window

ChatGPT has an average context window of around 128,000 tokens per conversation. That sounds enormous, but in practice it translates to roughly 96,000 words of conversation before ChatGPT starts forgetting the beginning of your chat.

Claude’s context window

Claude has a context window of around 200,000 tokens, making it one of the largest available right now. This is one of the reasons Claude is particularly strong for long document analysis and extended research tasks.

Gemini’s context window

Google’s Gemini 1.5 Pro offers a context window of up to 1 million tokens, currently the largest of any widely available model. This makes it particularly useful for analysing very long documents or large codebases.

Microsoft Copilot context window

Microsoft 365 Copilot uses GPT-4o with a context window of around 128,000 tokens, similar to ChatGPT Plus. For deep research tasks, the Researcher agent powered by Claude extends this significantly.

Why AI chatbot memory limits matter for your work

You get contradictory outputs

Once the AI forgets the beginning of your conversation, it loses the context you set up at the start. Instructions, tone, constraints and background information all disappear. The AI starts working from a much smaller picture and the outputs reflect that.

You get generic responses

Without the context you built up earlier, the AI falls back on generic patterns. The personalised, specific outputs you were getting at the start of the conversation become bland and unhelpful.

You waste time rebuilding context

If you do not know this is happening, you spend time trying to correct the AI, rephrasing prompts and wondering what went wrong. Understanding context windows saves you this frustration entirely.

How to fix AI chatbot memory limits: the practical solutions

These are the fixes that actually work, whether you are using ChatGPT, Claude, Copilot or Gemini.

Fix one: summarise the conversation strategically

When you notice outputs getting worse, do not start a new chat yet. First, ask the AI to summarise the conversation so far. Use a prompt like this:

“Summarise our conversation so far, including the key context, instructions and decisions we have made. Keep it concise but complete.”

Save that summary. You will use it in the next step.

Fix two: start a fresh chat with a contextual prompt

Take the summary from fix one and paste it into a brand new conversation as your opening prompt. Add any additional context the summary missed. This resets your token count while preserving everything important from the previous conversation.

A good opening prompt looks like this:

“Here is the context for our conversation. [Paste summary]. Based on this, please continue with [your next request].”

This is the single most effective way to work around context window limits without losing momentum.

Fix three: front load your context at the start of every conversation

Do not build context gradually. Put everything the AI needs to know at the very beginning of the conversation. Include your role, the task, the audience, the constraints and any relevant background. This maximises the useful life of your context window.

Fix four: keep conversations focused on one task

The more topics you cover in a single conversation, the faster you burn through your context window. Start a new conversation for each distinct task. This keeps your context window fresh and your outputs sharp.

Fix five: use Claude for long conversations

If you regularly hit context limits with ChatGPT, switch to Claude for extended tasks. Its 200,000 token window gives you significantly more runway before memory becomes an issue.

How AI chatbot memory limits affect different use cases

Long document analysis

If you are analysing a long report or contract, paste the document in first and ask your questions immediately. Do not build up a long conversation before introducing the document or you will burn through your context window before you get to the important work.

Multi-step projects

For projects that span multiple sessions, keep a running summary document outside the AI. Update it after each session and paste it in at the start of the next one.

Interview and CV preparation

If you are using AI to prepare for interviews or tailor your CV, keep your sessions focused. One session for CV tailoring, one for interview questions, one for company research. See our guide on how to show you are AI fluent on your CV for more on using AI in your job search.

Learning and upskilling

If you are using AI to learn new skills, shorter focused sessions work better than long rambling conversations. For structured AI learning with expert guidance, join ivee’s free sessions and courses.

How to get better at using AI chatbots in the UK

Understanding context windows is one of the most important things you can learn about working with AI. But it is just the start. The professionals who stand out in 2026 are the ones who understand how AI actually works, not just how to type a prompt.

Learn the fundamentals of how AI works

Understanding tokens, context windows, model differences and prompting patterns gives you a significant advantage over colleagues who are guessing. See ivee’s up to date prompting advice for AI in 2026 for the core patterns.

Join ivee’s free AI sessions and courses

ivee runs free sessions and courses covering everything from prompt engineering to AI agents, designed for UK professionals at every level. Join ivee’s AI Masterclass series to learn hands-on with expert guidance.

Prove your AI skills to employers

Once you understand how AI works, you can show employers you use it strategically, not just occasionally. See how to show you are AI fluent on your CV and get a free CV review for personalised feedback.

Conclusion

Your AI chatbot is not broken. It just has a memory limit. Once you understand context windows and how to work around them, you will get consistently better outputs and stop wasting time wondering what went wrong. The fix is simple. Summarise strategically, start fresh with context and keep your conversations focused. Master this and you are already ahead of most people using AI at work.

Further reading and sources

Want to understand AI well enough to actually get results from it?

Join ivee's free AI sessions and courses to learn how AI works, how to prompt it properly and how to use it to get ahead in your career.

Sign up for free

FAQs: what are AI chatbot memory limits?

Why does my AI chatbot forget what I said earlier?

Your AI chatbot has a context window, which is a memory limit measured in tokens. Once you hit the limit, the AI forgets the beginning of your conversation and starts working only from what it can still see.

This is why outputs get worse mid-conversation.

How many words can ChatGPT remember in one conversation?

ChatGPT has a context window of around 128,000 tokens, which translates to roughly 96,000 words. Once you exceed this, ChatGPT starts forgetting the earliest parts of your conversation.

Which AI chatbot has the biggest memory?

Google’s Gemini 1.5 Pro currently offers the largest context window at up to 1 million tokens. Claude has a context window of 200,000 tokens, making it one of the strongest options for long document analysis and extended conversations.

How do I stop my AI chatbot from forgetting my instructions?

Ask the AI to summarise the conversation before you hit the limit, then paste that summary into a new chat as your opening prompt.

Front loading context at the start of every conversation also helps you get more out of each session.

Where can I learn more about using AI effectively in the UK?

Join ivee’s free AI sessions and courses to learn prompt engineering, context management and AI workflows designed for UK professionals.

Visit ivee’s AI Masterclass series to get started with expert guidance.

Don't know what you don't know? Book a call.

Book a call and tell us where you're at. We'll show you how other teams are tackling AI, and, crucially, what's actually paying off.

Don't know what you don't know? Book a call.

Book a call and tell us where you're at. We'll show you how other teams are tackling AI, and, crucially, what's actually paying off.

Don't know what you don't know? Book a call.

Book a call and tell us where you're at. We'll show you how other teams are tackling AI, and, crucially, what's actually paying off.