I know this topic is not new when it comes to PlexAmp, Tidal, and music libraries already. But I was interested in whether a large language model such as ChatGPT 4 could help me find and surface films from my movie libraries.
ChatGPT 4 told me that my list of 5000 or so films would be no trouble to ingest via its file upload feature, which I did. (btw, simply producing a list of one’s library contents in Plex for export of any kind is conspicuously, and embarrassingly, absent from both client and server applications.)
I wanted help, for example, knowing whether I had films in my library that matched whatever collection idea I had in mind. But it can be difficult to identify all or most of the films that might match a collection such as “Swashbucklers” or films about grieving, or films that take place in submarines, or whatever. But a LLM can understand these types of requests easily.
So could a LLM such as ChatGPT 4 peruse a large list of film titles and pick out collections such as these? The short answer is, no.
ChatGPT is very open about its limitations here. By giving it a list of my film collection, it could only analyze the words in the list and match them against my queries much like a search engine might do. It cannot treat the contents of the list as stepping stones for cross reference into their plots, settings, etc. My submarine query, for example, pulled only 6 matches, including the obvious ones you’re thinking about like YELLOW SUBMARINE and RED OCTOBER, THE BOAT, K-19. But it said it matched these based on their titles only and could not look at the plot of, say, BLACK SEA, to determine it was also set entirely within a submarine. It could only “perceive” that when I asked about it specifically. Just as it told me THE WIZARD OF OZ would not match my query.
I see this as a temporary situation considering the rapid growth and investment in this field and in this product specifically. I’m writing this post to start a discussion and to start people thinking on the ways that an “AI”, so-to-speak, can help curate a large collection. I think most of us would be overjoyed to be able to feed plain language queries into a helper that can surface content both for consumption and for presentation. I’m hoping some devs are thinking about it, too.
Would love to hear further thoughts.