Comment on [deleted]
popcar2@piefed.ca 1 week ago
I like the idea of having as much info about a game as possible on one page. That said, I’m worried about how the data is being aggregated and how it won’t turn into a mess with more and more data being added. How are you going to stop bad info or duplicates from making it in? What happens when two sources report conflicting info?
At the end of a game page, it says “data aggregated from 5 sources”, but it should also list the 5 sources here.
Also, the page content doesn’t seem to be licensed. Most wikis and databases have a small blurb on how you can use the content. Terraria’s wiki for example says this at the bottom of every page:
Page content is under Creative Commons Attribution-NonCommercial-ShareAlike 4.0 License unless otherwise noted.
It would be nice if the page content can be explicitly licensed as creative commons.
Finally, this is just my opinion, but the website is a little to “what’s up fellow gamers” for me. The about page is kinda eye-rolling, it’d be nice to know what this website offers without all the “games for gamers, make your gaming identity now, gamer” talk.
faithcure@lemmy.world 1 week ago
Great questions, let me break down how the pipeline works:
Our first layer is deterministic — we pull core data from free APIs. Then multiple layers of cross-querying happen on the backend, and the results get merged before anything hits the site. After that, an AI agent runs research across the web (always with source attribution) and fills in the highest-confidence data.
On top of that, there’s a report feature so users can flag incorrect info. Honestly, I’d argue our data quality is already higher than most sites in this space because of this layered approach.
That said, the system is still in testing — nothing here is final, and it’ll settle into place and improve over time.
On the “5 sources” point: which piece of info came from where is actually indicated as subtext on the page. But I hear you that it could be surfaced more clearly.
On sourcing and licensing — we’re strict about this. We only use free APIs. For example, we don’t pull any data from MobyGames because their terms are restrictive; we only do read-only matching to check whether existing data lines up.
As for licensing our own page content (the Terraria wiki example) — that’s a really good catch, and it’s not something we’ve formalized yet. Adding it to the to-do list. Genuinely valuable feedback, thank you.
popcar2@piefed.ca 1 week ago
Okay this is definitely AI spam lol
faithcure@lemmy.world 1 week ago
Not spam. I’m just using translation for best communication. Thank you.