Data AnalysisVerified
Reviewed and published by trentmaziarz, April 21, 2026. Discovered and drafted by our automated research pipeline.
Artsy uses an automated data pipeline to compute an artwork similarity graph that underpins The Art Genome Project's discovery features, processing large batches of artwork data offline to generate similarity scores used in search and browse results.
Details
According to Artsy's engineering blog, the artwork similarity graph that powers The Art Genome Project is processed offline by a generic job engine written in Ruby or by Amazon Elastic MapReduce. The system takes data snapshots from MongoDB, runs computation jobs, and exports results back to the production database. Artsy also uses Jupyter Notebooks with pandas and scikit-learn for more in-depth data analysis work. This pipeline feeds the similarity scores that surface related artworks throughout the platform.
Products affected
Artsy web platformArtsy iOS app
Sources & Evidence
Company Disclosure
Cite this record
Trace Foundation. (2026). Artsy: Artsy uses an automated data pipeline to compute an artwork similarity graph that underpins The Art Genome Project's discovery features, processing large batches of artwork data offline to generate similarity scores used in search and browse results (data as of 2026-04-21) [Data set record]. AI Trace. https://www.aitrace.org/r/practice/ac29b7a9-f56b-4cbf-b1a7-5fb8c5db25d7. Accessed September 12, 2026.
- Stable link
- https://www.aitrace.org/r/practice/ac29b7a9-f56b-4cbf-b1a7-5fb8c5db25d7
- Data as of
- April 21, 2026
- Last verified
- April 21, 2026
Other practices by Artsy
Have evidence about Artsy's AI practices? Submit a report.
Report a Sighting →