This is Part 3 and the final post in a three-part series on why long-context language models still struggle with memory.
In Part 1, we saw why increasing context length does not equal better memory.
In Part 2, we reframed sequence models as memory systems using the MIRAS perspective.
In this final post, we examine Titans, a concrete architecture...