Catalog Normalizer — ES/EN product feeds (MXN/USD) avatar

Catalog Normalizer — ES/EN product feeds (MXN/USD)

Pricing

from $0.50 / 1,000 normalized products

Go to Apify Store
Catalog Normalizer — ES/EN product feeds (MXN/USD)

Catalog Normalizer — ES/EN product feeds (MXN/USD)

Clean messy Spanish/English product catalogs: titles, brands, price+currency (MXN/USD), stock, attributes. For MercadoLibre, Shopify, B2B feeds and AI agents.

Pricing

from $0.50 / 1,000 normalized products

Rating

0.0

(0)

Developer

Armando Cortés

Armando Cortés

Maintained by Community

Actor stats

0

Bookmarked

2

Total users

1

Monthly active users

14 hours ago

Last modified

Categories

Share

MX Catalog Normalizer

Normalize messy product catalogs into a consistent ecommerce-ready dataset for Mexico and international workflows. The Actor accepts product objects in Spanish or English and returns standardized titles, brands, prices, currencies, availability, attributes, and quality anomalies.

What it does

  • Maps Spanish and English aliases such as nombre/name, marca/brand, and precio/price.
  • Parses common MXN, USD, and EUR price formats.
  • Standardizes availability to in_stock, out_of_stock, preorder, or unknown.
  • Preserves unknown source fields inside attributes.
  • Flags missing or invalid values instead of inventing data.
  • Processes up to 10,000 products per run.
  • Makes no network requests and uses no paid third-party APIs.

Input

{
"products": [
{
"nombre": "Laptop Profesional 14 pulgadas",
"marca": "EjemploMX",
"precio": "$18,499.90 MXN",
"availability": "disponible",
"sku": "DEMO-001"
}
]
}

Output

{
"title": "Laptop Profesional 14 pulgadas",
"brand": "EjemploMX",
"price": 18499.9,
"currency": "MXN",
"availability": "in_stock",
"attributes": {
"sku": "DEMO-001"
},
"anomalies": [],
"normalized_at": "2026-09-30T01:09:42.935Z"
}

Results are written to the run's default Dataset and can be downloaded as JSON, CSV, Excel, XML, RSS, or JSONL using Apify's standard exports.

Limits

  • Maximum 10,000 products per run.
  • Maximum 64 KiB serialized per product.
  • Maximum nesting depth: 6.
  • Maximum 80 keys per product.
  • Maximum string length: 4,096 characters.
  • Maximum attributes size: 32 KiB.

The Actor rejects oversized or malformed payloads before copying attributes or writing results.

Data quality behavior

This Actor normalizes supplied values; it does not verify whether commercial claims are true. Missing fields remain null and appear in anomalies. Review results before importing them into a production store or ERP.

Privacy

Do not submit personal data, credentials, payment details, or confidential information. The Actor does not transmit input to third parties. Input and Dataset retention follow your Apify account and storage settings; delete storages when no longer needed.

Performance and security

The verified benchmark processed 1,000 items across 20 runs with mean 9.242 ms and approximately 108,205 items/s on the test host. Actual cloud runtime includes container startup and Dataset writes.

Verified release:

  • Local tests: 7/7 passed.
  • Apify cloud build: successful.
  • Remote smoke run: successful.
  • Production dependency audit: 0 known vulnerabilities.
  • Runtime permission level: LIMITED_PERMISSIONS.

Español

Normaliza catálogos heterogéneos para ecommerce en México. Acepta campos en español o inglés, estandariza títulos, marcas, precios, monedas y disponibilidad, conserva atributos adicionales y reporta anomalías sin inventar información.

La entrada es {"products":[...]} y el límite es de 10,000 productos por ejecución. Los resultados quedan disponibles en el Dataset de Apify para descargarse en JSON, CSV, Excel, XML, RSS o JSONL.

No envíes datos personales, credenciales ni información confidencial. El Actor no hace scraping, no llama servicios externos y no usa APIs pagadas.

Support

For reproducible bugs, include a redacted sample input, the run ID, expected output, and actual output. Never include API tokens or sensitive customer data.