Instance view of the static supplierName (keeps this.supplierName working).
Instance view of the static baseURL (keeps this.baseURL working).
PrivatesupplierThis supplier's own class, typed for static metadata access. this.constructor
is otherwise typed Function (no custom statics), so the instance getters below
read the concrete class's static fields through this narrowed view.
ProtectedsupportsInstance view of the static supportsCAS flag (keeps this.supportsCAS working).
ProtectedsupportsInstance view of the static supportsFormula flag.
ProtectedsupportsInstance view of the static supportsSMILES flag.
Instance view of the static shipping scope (keeps this.shipping working).
Instance view of the static country (keeps this.country working).
Instance view of the static paymentMethods (keeps this.paymentMethods working).
ProtectedshipsInstance view of the static shipsTo allowlist (keeps this.shipsTo working).
ProtectedapiInstance view of the static apiURL (keeps this.apiURL working).
StaticrequiredAll host origin patterns required for this supplier to function.
Automatically includes baseURL and, if defined, apiURL. Used by the
factory to check chrome permissions before querying.
Instance view of the static requiredHosts.
ProtectedeffectiveThe single query term this supplier should search by. Normally the raw query, but when the query is a CAS/formula/SMILES identifier the supplier can't search for itself (per supportsCAS/supportsFormula/supportsSMILES) and the factory resolved it, it's the broadest resolved name — the one likely to yield the most store results — so a name-only supplier can answer "Na6O18P6" by searching "Sodium hexametaphosphate". Matching then scores results against every candidate (see effectiveQueryCandidates/fuzzyScore). Falls back to the raw query when the type is supported, the query is a plain name, or nothing was resolved.
The effective search query.
// this.query === "10124-56-8", supportsCAS === false, resolved names available
this.effectiveQuery; // "Sodium hexametaphosphate"
Floors the results limit at 15. ScienceLab is a single-store catalog search (one sitemap fetch), so surfacing more results than the global default (5) is cheap and useful; an explicit larger caller limit is still honored.
The search query.
Optionallimit: numberThe requested results limit (raised to at least 15).
Optionalcontroller: AbortControllerOptional abort controller for the search.
new SupplierScienceLab('acetone').limit; // 15
Determines whether this supplier ships to the given country. Prefers the
explicit shipsTo allowlist when the supplier declares one; otherwise falls
back to the coarse shipping scope — "worldwide"/"international" ship
anywhere, while "domestic"/"local" ship only within the supplier's own
country.
The user's location as an ISO 3166-1 alpha-2 country code.
True if the supplier ships to location, false otherwise.
// Supplier with shipsTo = ["US", "CA"]:
supplier.shipsToCountry("US"); // true
supplier.shipsToCountry("DE"); // false
// Domestic US supplier (no shipsTo):
supplier.shipsToCountry("US"); // true
supplier.shipsToCountry("DE"); // false
// Worldwide supplier (no shipsTo):
supplier.shipsToCountry("DE"); // true
public shipsToCountry(location: CountryCode): boolean {
// Pass an explicit meta object rather than `this.constructor` so the
// (protected) `shipsTo` static doesn't clash with SupplierStaticMeta's public
// shape; the getters read the concrete class's static values.
return SupplierBase.shipsToCountryStatic(
{ shipping: this.shipping, country: this.country, shipsTo: this.shipsTo },
location,
);
}
StaticshipsWhether a supplier with the given static shipping metadata ships to
location. Shared by the instance shipsToCountry and SupplierFactory
so the UI can test shipping compatibility from a supplier's static fields
without instantiating it.
The supplier's static shipping/country/shipsTo metadata.
Destination country (ISO 3166-1 alpha-2).
True when the supplier ships to location.
public static shipsToCountryStatic(meta: SupplierStaticMeta, location: CountryCode): boolean {
if (meta.shipsTo) {
return meta.shipsTo.includes(location);
}
switch (meta.shipping) {
case 'worldwide':
case 'international':
return true;
case 'domestic':
case 'local':
return meta.country === location;
default:
return true;
}
}
Initializes the cache for the supplier. This is called after construction to ensure supplierName is set.
The cache is initialized with the supplier's name and is used to store both query results and product data. This method should be called after the supplier's name is set to ensure proper cache key generation.
class MySupplier extends SupplierBase<Product> {
constructor() {
super("acetone", 5);
// supplierName is set here
this.initCache(); // Initialize cache after supplierName is set
}
}
public initCache(
enabled: boolean = true,
doNotCacheEmptyResults: boolean = false,
cacheTtlMinutes: number = 0,
noCacheStatusCodes: number[] = [429],
): void {
this.cache = new SupplierCache(
this.supplierName,
this.constructor.name,
enabled,
doNotCacheEmptyResults,
cacheTtlMinutes,
);
// Stored on the supplier (not the cache): the decision is made at cache-write time in
// getProductData(WithCache), where the per-product fetch status is known.
this.noCacheStatusCodes = noCacheStatusCodes ?? [429];
}
Applies (or clears) a runtime override for the fuzz scorer. Driven by
userSettings.fuzzScorerOverride — when the user picks a scorer in the
Advanced drawer section, SupplierFactory calls this on each instance
so the choice takes effect uniformly across every supplier.
Silently ignores unknown names so an outdated / corrupted setting can't blow up the search flow — callers fall back to the subclass default.
Name of a scorer from FUZZ_SCORERS, or undefined to
clear the override and use the subclass default.
const supplier = new MySupplier("acetone", 5, controller);
supplier.setFuzzScorerOverride("token_set_ratio");
// fuzzyFilter now uses token_set_ratio regardless of MySupplier's default
supplier.setFuzzScorerOverride(undefined);
// back to MySupplier's default
public setFuzzScorerOverride(name: string | undefined): void {
console.debug('setFuzzScorerOverride', { name });
if (isFuzzScorerName(name)) {
this.fuzzScorerOverride = FUZZ_SCORERS[name];
} else {
this.fuzzScorerOverride = undefined;
}
}
Applies a runtime override for maxAllowableSearchTimeSec, driven by
userSettings.maxAllowableSearchTimeSec (set in the Advanced settings section). Accepts the raw
setting value (which may arrive as a string from the number input). An absent, empty, or
invalid value is ignored so the supplier keeps its class default; a valid non-negative number
(including 0 to disable the limit) replaces it.
The override in seconds, or any value (invalid/empty input is ignored)
supplier.setMaxAllowableSearchTimeSec(60); // cap searches at 60s
supplier.setMaxAllowableSearchTimeSec(""); // no-op, keep the per-supplier default
public setMaxAllowableSearchTimeSec(value: unknown): void {
if (value === undefined || value === null || value === '') {
return;
}
const seconds = Number(value);
if (!Number.isNaN(seconds) && seconds >= 0) {
this.maxAllowableSearchTimeSec = seconds;
}
}
Sets the parsed advanced-search query for this instance. Called by
SupplierFactory once per search so every supplier shares the same parse of
the user's input.
The parsed query, or undefined to clear it.
supplier.setParsedQuery(parseSearchQuery("Sodium OR Potassium"));
public setParsedQuery(parsed: ParsedSearchQuery | undefined): void {
this.parsedQuery = parsed;
}
Applies a runtime override for fuzzyFilteringDisabled, driven by
userSettings.fuzzyFilteringDisabled. When true, fuzzball scoring is skipped
and only the boolean predicate (substring matching) is applied.
True to disable fuzzy filtering, false (default) to keep it.
supplier.setFuzzyFilteringDisabled(true); // show raw/boolean-only results
public setFuzzyFilteringDisabled(value: boolean): void {
this.fuzzyFilteringDisabled = value === true;
}
Sets the map of resolved structure terms for this instance. Called by
SupplierFactory once per search so every supplier shares one resolution of
any SMILES/structure terms instead of each hitting the network.
Map of raw search term → resolved structure, or undefined when none.
supplier.setResolvedStructures(new Map([["CCO", { name: "ethanol", cas: ["64-17-5"] }]]));
public setResolvedStructures(resolved: ReadonlyMap<string, ResolvedStructure> | undefined): void {
this.resolvedStructures = resolved;
this.effectiveQueryCandidatesCache = undefined;
}
ProtectedeffectiveThe candidate queries this supplier matches against. Normally just
[this.query], but when the query is a CAS/formula/SMILES identifier the
supplier can't search itself (per supportsCAS/supportsFormula/
supportsSMILES), it's the factory-resolved chemical-name candidates —
the compound Title plus cleaned synonyms — so a product can be matched by
whichever name scores best (see fuzzyScore). Falls back to
[this.query] when the type is supported, the query is a plain name, or
nothing was resolved. Memoized.
The candidate queries, best first (never empty).
// this.query === "10124-56-8", supportsCAS === false, resolution available
this.effectiveQueryCandidates(); // ["Hexasodium hexametaphosphate", "Sodium hexametaphosphate", …]
protected effectiveQueryCandidates(): string[] {
if (this.effectiveQueryCandidatesCache === undefined) {
this.effectiveQueryCandidatesCache = this.resolveEffectiveQueryCandidates();
}
return this.effectiveQueryCandidatesCache;
}
ProtectedgetReturns the parsed query for this instance. Uses the effectiveQuery, so
an identifier query the supplier can't search is parsed as its resolved name.
Otherwise lazily parses this.query when SupplierFactory did not set a
parsed query (e.g. in unit tests that construct a supplier directly).
The parsed search query.
const { isAdvanced, ast } = this.getAst();
protected getAst(): ParsedSearchQuery {
const effective = this.effectiveQuery;
if (effective !== this.query) {
return parseSearchQuery(effective);
}
return this.parsedQuery ?? parseSearchQuery(this.query);
}
ProtectedsetupPlaceholder for any setup that needs to be done before the query is made. Override this in subclasses if you need to perform setup (e.g., authentication, token fetching).
A promise that resolves when the setup is complete.
await supplier.setup();
protected async setup(): Promise<void> {}
ProtectedhttpRetrieves HTTP headers from a URL using a HEAD request. Useful for checking content types, caching headers, and other metadata without downloading the full response.
The URL to fetch headers from
Promise resolving to the response headers or void if request fails
// Basic usage
const headers = await supplier.httpGetHeaders('https://example.com/product/123');
if (headers) {
console.log('Content-Type:', headers['content-type']);
}
// With error handling
try {
const headers = await supplier.httpGetHeaders('https://example.com/product/123');
if (headers) {
console.log('Headers:', headers);
}
} catch (err) {
console.error('Failed to fetch headers:', err);
}
protected async httpGetHeaders(url: string | URL): Promise<Maybe<HeadersInit>> {
const requestObj = new Request(this.href(url), {
signal: this.controller.signal,
headers: new Headers(this.headers),
referrer: this.baseURL,
referrerPolicy: 'strict-origin-when-cross-origin',
body: null,
method: 'HEAD',
mode: 'cors',
credentials: 'include',
});
try {
const httpResponse = await this.fetch(requestObj);
return Object.fromEntries(httpResponse.headers.entries()) satisfies HeadersInit;
} catch (error: unknown) {
if (error instanceof Error && error.name === 'AbortError') {
this.logger.warn('Request was aborted', { error, signal: this.controller.signal });
this.controller.abort('Abort signal detected');
} else {
this.logger.error('Error received during fetch:', {
error,
signal: this.controller.signal,
});
}
return;
}
}
ProtectedhttpSends a POST request to the given URL with the given body and headers. Handles request setup, error handling, and response caching.
The request configuration options
Promise resolving to the Response object or void if request fails
// Basic POST request
const response = await supplier.httpPost({
path: '/api/v1/products',
body: { name: 'Test Chemical' }
});
// POST with custom headers
const response = await supplier.httpPost({
path: '/api/v1/products',
body: { name: 'Test Chemical' },
headers: {
'Authorization': 'Bearer token123',
'Content-Type': 'application/json'
}
});
// POST with custom host and params
const response = await supplier.httpPost({
path: '/api/v1/products',
host: 'api.example.com',
body: { name: 'Test Chemical' },
params: { version: '2' }
});
// Error handling
try {
const response = await supplier.httpPost({ path: '/api/v1/products', body: { name: 'Test' } });
if (response && response.ok) {
const data = await response.json();
console.log('Created:', data);
}
} catch (err) {
console.error('POST failed:', err);
}
protected async httpPost({
path,
host,
body,
params,
headers,
}: RequestOptions): Promise<Maybe<Response>> {
this.logger.log('httpPost| Requesting:', {
path,
host,
body,
params,
headers,
});
const method = 'POST';
const mode = 'cors';
const referrer = this.baseURL;
const referrerPolicy = 'strict-origin-when-cross-origin';
const signal = this.controller.signal;
const headersObj = new Headers({
...this.headers,
...headers,
});
let bodyStr = null;
if (body instanceof FormData) {
headersObj.set('Content-Type', 'application/x-www-form-urlencoded; charset=UTF-8');
bodyStr = body;
} else if (typeof body === 'string') {
bodyStr = body;
} else if (typeof body === 'object' && body !== null) {
bodyStr = JSON.stringify(body);
}
const url = this.href(path, params, host);
const requestObj = new Request(url, {
signal,
headers: headersObj,
referrer,
referrerPolicy,
body: bodyStr,
method,
mode,
credentials: 'include',
});
// Fetch the goods
const httpResponse = await this.fetch(requestObj);
if (!isHttpResponse(httpResponse) || !httpResponse.ok) {
const badResponse = await httpResponse.text();
this.logger.error('Invalid POST response: ', badResponse);
throw new TypeError(`Invalid POST response: ${String(httpResponse)}`);
}
return httpResponse;
}
ProtectedhttpSends a POST request with the body encoded as multipart/form-data.
Converts the given object into a FormData instance (one field per
key/value pair) before delegating to httpPost.
The request configuration options. body must be a non-null object.
Promise resolving to the Response object or void if the request fails
TypeError - If body is not an object, or the response is not a valid HTTP response
const response = await this.httpPostFormData({
path: "/api/v1/cart",
body: { productId: "123", quantity: "2" },
});
protected async httpPostFormData({
path,
host,
body,
params,
headers,
}: RequestOptions): Promise<Maybe<Response>> {
if (typeof body !== 'object' || body === null) {
throw new TypeError('httpPostFormData| Body must be an object');
}
headers = {
...headers,
'Content-Type': 'application/x-www-form-urlencoded',
};
const formData = new FormData();
for (const [key, value] of Object.entries(body)) {
formData.append(key, value);
}
const httpResponse = await this.httpPost({ path, host, body: formData, params, headers });
if (!isHttpResponse(httpResponse) || !httpResponse.ok) {
const badResponse = await httpResponse?.text();
this.logger.error('Invalid POST response: ', badResponse);
throw new TypeError(`Invalid POST response: ${String(httpResponse)}`);
}
this.logger.log('httpPostFormData| Successfully sent POST request to:', path);
return httpResponse;
}
ProtectedhttpSends a POST request and returns the response as a JSON object.
The parameters for the POST request.
The response from the POST request as a JSON object.
// Basic usage
const data = await supplier.httpPostJson({
path: '/api/v1/products',
body: { name: 'John' }
});
// With custom headers and error handling
try {
const data = await supplier.httpPostJson({
path: '/api/v1/products',
body: { name: 'John' },
headers: { 'Authorization': 'Bearer token123' }
});
if (data) {
console.log('Created:', data);
}
} catch (err) {
console.error('POST JSON failed:', err);
}
protected async httpPostJson({
path,
host,
body,
params,
headers,
}: RequestOptions): Promise<Maybe<JsonValue>> {
const httpResponse = await this.httpPost({ path, host, body, params, headers });
if (!isJsonResponse(httpResponse) || !httpResponse.ok) {
this.logger.error('httpPostJson| Invalid POST response: ', {
httpResponse,
path,
host,
body,
params,
headers,
});
throw new TypeError(`httpPostJson| Invalid POST response: ${httpResponse}`);
}
return await httpResponse.json();
}
ProtectedhttpSends a POST request and returns the response as a HTML string.
The request configuration options
Promise resolving to the HTML response as a string or void if request fails
TypeError - If the response is not valid HTML content
// Basic usage
const html = await supplier.httpPostHtml({
path: '/api/v1/products',
body: { name: 'John' }
});
protected async httpPostHtml({
path,
host,
body,
params,
headers,
}: RequestOptions): Promise<Maybe<string>> {
const httpResponse = await this.httpPost({ path, host, body, params, headers });
if (!isHtmlResponse(httpResponse)) {
throw new TypeError(`httpPostHtml| Invalid POST response: ${httpResponse}`);
}
return await httpResponse.text();
}
ProtectedhttpSends a GET request to the given URL with the specified options. Handles request setup, error handling, and response caching.
The request configuration options
Promise resolving to the Response object or void if request fails
// Basic GET request
const response = await supplier.httpGet({
path: '/products/search',
params: { query: 'sodium chloride' }
});
// GET with custom headers
const response = await supplier.httpGet({
path: '/products/search',
headers: { 'Accept': 'application/json' }
});
// GET with custom host
const response = await supplier.httpGet({
path: '/products/search',
host: 'api.example.com',
params: { category: 'chemicals' }
});
// Error handling
try {
const response = await supplier.httpGet({ path: '/products/search' });
if (response && response.ok) {
const data = await response.json();
console.log('Products:', data);
}
} catch (err) {
console.error('GET failed:', err);
}
protected async httpGet({
path,
params,
headers,
host,
rethrowErrors,
}: RequestOptions): Promise<Maybe<Response>> {
// Check if the request has been aborted before proceeding
if (this.controller.signal.aborted) {
this.logger.warn('Request was aborted before fetch', {
signal: this.controller.signal,
});
return;
}
const headersRaw = { ...this.headers };
Object.assign(headersRaw, {
accept: [
'text/html',
'application/xhtml+xml',
'application/xml;q=0.9',
'image/avif',
'image/webp',
'image/apng',
'*/*;q=0.8',
].join(','),
...(headers ?? {}),
});
const requestObj = new Request(this.href(path, params, host), {
signal: this.controller.signal,
headers: new Headers(headersRaw),
referrer: this.baseURL,
referrerPolicy: 'no-referrer',
body: null,
method: 'GET',
mode: 'cors',
credentials: 'include',
redirect: 'follow',
});
try {
// Fetch the goods
const httpResponse = await this.fetch(requestObj.url, requestObj);
const responseHeaders = Object.fromEntries(
httpResponse.headers.entries(),
) satisfies HeadersInit;
this.logger.debug('responseHeaders:', responseHeaders);
this.logger.debug('responseHeaders.location:', responseHeaders.location);
return httpResponse;
} catch (error: unknown) {
if (error instanceof Error && error.name === 'AbortError') {
this.logger.warn('Request was aborted', { error, signal: this.controller.signal });
this.controller.abort('Abort signal detected');
return;
}
this.logger.error('Error received during fetch:', {
error,
signal: this.controller.signal,
});
// Opt-in: surface the failure (e.g. an HttpError 429) so the caller can apply
// status-aware retry/backoff. Default behavior remains swallow-and-return-undefined.
if (rethrowErrors) {
throw error;
}
return;
}
}
ProtectedfuzzyScores a single string against the current search query (this.query)
using the active fuzz scorer — the user's fuzzScorerOverride when set,
otherwise the supplier's fuzzScorer. Returns a 0–100 similarity score
(higher is closer). Useful for suppliers that can only fuzz-match after a
secondary request reveals the real product name, e.g. when the search
index only exposes coarse category breadcrumbs.
The text to score against this.query.
A similarity score from 0 (no match) to 100 (identical).
// this.query === "sodium borohydride"
this.fuzzyScore("Sodium borohydride, min 95%"); // ~90
this.fuzzyScore("Acetone"); // ~10
protected fuzzyScore(text: string): number {
const activeScorer = this.fuzzScorerOverride ?? this.fuzzScorer;
// Score against every effective-query candidate and keep the best — for an
// identifier query these are the resolved name + synonyms, so a product
// matches whichever name it names. A plain query has a single candidate.
let best = 0;
for (const candidate of this.effectiveQueryCandidates()) {
const score = activeScorer(candidate, text);
if (score > best) {
best = score;
}
}
return best;
}
ProtectedfuzzyFilters an array of data using fuzzy string matching to find items that closely match a query string. Uses the WRatio algorithm from fuzzball for string similarity comparison.
The search string to match against
Array of data objects to search through
Minimum match percentage (0-100) for a match to be included (default: 55)
Array of matching data objects with added fuzzy match metadata
// Example with simple string array
const products = [
{ title: "Sodium Chloride", price: 29.99 },
{ title: "Sodium Hydroxide", price: 39.99 },
{ title: "Potassium Chloride", price: 19.99 }
];
const matches = this.fuzzyFilter("sodium chloride", products);
// Returns: [
// {
// title: "Sodium Chloride",
// price: 29.99,
// _fuzz: { score: 100, idx: 0 }
// },
// {
// title: "Sodium Hydroxide",
// price: 39.99,
// _fuzz: { score: 85, idx: 1 }
// }
// ]
// Example with custom minMatchPercentage
const strictMatches = this.fuzzyFilter("sodium chloride", products, 90);
// Returns only exact matches with score >= 90
// Example with different data structure
const chemicals = [
{ name: "NaCl", formula: "Sodium Chloride" },
{ name: "NaOH", formula: "Sodium Hydroxide" }
];
// Override titleSelector to use formula field
this.titleSelector = (data) => data.formula;
const formulaMatches = this.fuzzyFilter("sodium chloride", chemicals);
protected fuzzyFilter<X>(
query: string,
data: X[],
minMatchPercentage: number = this.minMatchPercentage,
): X[] {
// User's Advanced-settings override wins over the subclass default.
const activeScorer = this.fuzzScorerOverride ?? this.fuzzScorer;
// console.log(
// `[fuzzyFilter] ${this.supplierName} query="${query}" — scorer comparison (cutoff=${minMatchPercentage})`,
// );
if (IS_DEV_BUILD) {
this.showFuzzScorerComparisonTable(query, data);
}
if (this.fuzzyFilterRankOnly) {
// Rank every candidate by score (no cutoff) and return them in score order; the
// caller slices the top N. Avoids dropping clear matches whose ratio-style score
// falls under minMatchPercentage purely because the title dwarfs the query.
return extract(query, data, {
scorer: activeScorer,
processor: this.titleSelector,
sortBySimilarity: true,
}).map(([obj, score, idx]) => this.attachFuzz(obj, score, idx));
}
const results = extract(query, data, {
scorer: activeScorer,
processor: this.titleSelector,
cutoff: minMatchPercentage,
sortBySimilarity: true,
}).reduce<FuzzyMatchResult<X>[]>((acc, [obj, score, idx]) => {
if (score < minMatchPercentage) {
this.logger.debug('fuzzyFilter: score below minimum match percentage, excluding product', {
product: obj,
score,
idx,
minMatchPercentage: minMatchPercentage,
});
return acc;
}
acc[idx] = Object.assign(obj, { _fuzz: { score, idx }, matchPercentage: score });
return acc;
}, []);
this.logger.debug('[fuzzyFilter]', {
supplierName: this.supplierName,
query,
minMatchPercentage,
activeScorer,
results,
});
// Get rid of any empty items that didn't match closely enough
return results.filter((item) => !!item);
}
ProtectedfuzzyAdvanced-search-aware companion to fuzzyFilter. Filters data using
the parsed query (getAst) so boolean operators (AND/OR/NOT) and
nesting are honored, and respects the fuzzyFilteringDisabled toggle.
Behavior matrix:
data unchanged (raw supplier results).Array of raw search-result objects to filter.
Minimum leaf match score when fuzzing is on.
The filtered (and, when fuzzing, ranked) subset, each item tagged
with _fuzz/matchPercentage like fuzzyFilter.
// this.query === "Sodium OR Potassium"
const matches = this.fuzzyFilterAst(products);
protected fuzzyFilterAst<X>(
data: X[],
minMatchPercentage: number = this.minMatchPercentage,
): X[] {
const parsed = this.getAst();
if (!parsed.isAdvanced) {
// Plain query: no filtering when disabled; the multi-candidate path for a
// resolved identifier query (several candidate names); else the legacy
// single-query fuzzy path.
if (this.fuzzyFilteringDisabled) {
return data;
}
if (this.effectiveQueryCandidates().length > 1) {
return this.fuzzyFilterCandidates(data, minMatchPercentage);
}
return this.fuzzyFilter(parsed.raw.trim(), data, minMatchPercentage);
}
const scorer = this.fuzzScorerOverride ?? this.fuzzScorer;
// Rank-only floors the leaf score at 0 so predicate-matching items are never dropped
// for a low fuzz score; the sort below still ranks them. Otherwise enforce the cutoff.
const threshold = this.fuzzyFilteringDisabled
? 1
: this.fuzzyFilterRankOnly
? 0
: minMatchPercentage;
const fuzzyWords = !this.fuzzyFilteringDisabled;
const matched = data.reduce<FuzzyMatchResult<X>[]>((acc, obj, idx) => {
const title = this.titleSelector(obj) ?? '';
const score = scoreAstMatch(title, parsed.ast, { scorer, threshold, fuzzyWords });
if (score === null) {
return acc;
}
acc.push(this.attachFuzz(obj, score, idx));
return acc;
}, []);
// Rank by relevance when fuzzing; preserve backend order when disabled.
if (!this.fuzzyFilteringDisabled) {
matched.sort((a, b) => (b.matchPercentage ?? 0) - (a.matchPercentage ?? 0));
}
return matched;
}
ProtectedfuzzyAdvanced-search-aware, keep-or-drop companion to fuzzyScore for
suppliers that re-filter a single title after a detail fetch (e.g. LiMac).
Returns the score to keep the item, or null to drop it, honoring both the
parsed query and the fuzzyFilteringDisabled toggle:
minMatchPercentage;
advanced query keeps items satisfying the predicate with fuzzy leaf scores.The text (e.g. a detail-page product name) to score.
A 0–100 score to keep the item, or null to drop it.
// this.query === "acid AND NOT boric"
this.fuzzyScoreAst("Sulfuric acid"); // a number
this.fuzzyScoreAst("Boric acid"); // null
protected fuzzyScoreAst(text: string): number | null {
const parsed = this.getAst();
if (!parsed.isAdvanced) {
if (this.fuzzyFilteringDisabled) {
return 100;
}
const score = this.fuzzyScore(text);
return score >= this.minMatchPercentage ? score : null;
}
const scorer = this.fuzzScorerOverride ?? this.fuzzScorer;
const threshold = this.fuzzyFilteringDisabled ? 1 : this.minMatchPercentage;
return scoreAstMatch(text, parsed.ast, {
scorer,
threshold,
fuzzyWords: !this.fuzzyFilteringDisabled,
});
}
ProtectedderiveDerives the backend search terms for a keyword-only supplier from an advanced query: one representative term per positive OR-group (the longest — most selective — token of each AND-group), de-duplicated and capped at maxFallbackQueries. Returns an empty array when there are no positive terms (e.g. a purely negative query), in which case the caller falls back to a single raw search.
The de-duplicated, capped list of backend search terms.
// this.query === "(Sodium OR Potassium) AND Hydroxide"
this.deriveFallbackTerms(); // ["Hydroxide", "Hydroxide"] -> ["Hydroxide"] (deduped)
protected deriveFallbackTerms(): string[] {
const groups = extractOrGroups(this.getAst().ast);
// How many AND-groups each term appears in — a term shared across groups (a
// common factor, e.g. "Hydroxide" in "(Sodium OR Potassium) AND Hydroxide")
// covers more of the query in fewer requests, so prefer it; break ties by
// length (more selective). Picking one representative term per group keeps
// each request's result set a superset of that group's matches.
const frequency = new Map<string, number>();
for (const group of groups) {
for (const term of new Set(group)) {
frequency.set(term, (frequency.get(term) ?? 0) + 1);
}
}
const terms = groups
.map(
(group) =>
group.slice().sort((a, b) => {
const byFrequency = (frequency.get(b) ?? 0) - (frequency.get(a) ?? 0);
return byFrequency !== 0 ? byFrequency : b.length - a.length;
})[0],
)
.filter((term): term is string => Boolean(term));
return [...new Set(terms)].slice(0, this.maxFallbackQueries);
}
ProtectedhttpMakes an HTTP GET request and returns the response as a string. Handles request configuration, error handling, and HTML parsing.
The request configuration options
Promise resolving to the HTML response as a string or void if request fails
TypeError - If the response is not valid HTML content
// Basic GET request
const html = await this.httpGetHtml({
path: "/api/products",
params: { search: "sodium" }
});
// GET request with custom headers
const html = await this.httpGetHtml({
path: "/api/products",
headers: {
"Authorization": "Bearer token123",
"Accept": "text/html"
}
});
// GET request with custom host
const html = await this.httpGetHtml({
path: "/products",
host: "api.supplier.com",
params: { limit: 10 }
});
protected async httpGetHtml({
path,
params,
headers,
host,
}: RequestOptions): Promise<Maybe<string>> {
const httpResponse = await this.httpGet({ path, params, headers, host });
if (!isHtmlResponse(httpResponse)) {
throw new TypeError(`httpGetHtml| Invalid GET response: ${httpResponse}`);
}
return await httpResponse.text();
}
ProtectedhttpMakes an HTTP GET request and returns the response as parsed JSON. Handles request configuration, error handling, and JSON parsing.
The request configuration options
Promise resolving to the parsed JSON response or void if request fails
TypeError - If the response is not valid JSON content
// Basic GET request
const data = await supplier.httpGetJson({ path: '/api/products', params: { search: 'sodium' } });
// GET request with custom headers
const data = await supplier.httpGetJson({
path: '/api/products',
headers: {
'Authorization': 'Bearer token123',
'Accept': 'application/json'
}
});
// GET request with custom host
const data = await supplier.httpGetJson({
path: '/products',
host: 'api.supplier.com',
params: { limit: 10 }
});
// Error handling
try {
const data = await supplier.httpGetJson({ path: '/api/products' });
if (data) {
console.log('Products:', data);
}
} catch (error) {
console.error('Failed to fetch products:', error);
}
protected async httpGetJson({
path,
params,
headers = {},
host,
}: RequestOptions): Promise<Maybe<JsonValue>> {
if (
typeof headers?.accept === 'undefined' ||
!Array.isArray(headers?.accept) ||
!headers?.accept.includes('application/json')
) {
headers.accept = ['application/json', 'text/plain', '*/*'].join(',');
}
const httpRequest = await this.httpGet({ path, params, headers, host });
if (!isJsonResponse(httpRequest)) {
const badResponse = isHttpResponse(httpRequest) ? await httpRequest.text() : undefined;
this.logger.error('Invalid HTTP GET JSON response:', {
badResponse,
httpRequest,
path,
params,
headers,
host,
});
return;
}
return await httpRequest.json();
}
ProtectedqueryExecutes a product search query with caching support. First checks the cache for existing results, then falls back to the actual query if needed. The limit parameter is only used for the actual query and doesn't affect caching.
The search term to query products for
The maximum number of results to return (defaults to instance limit)
Promise resolving to array of product builders or void if search fails
// Basic usage with default limit
const results = await supplier.queryProductsWithCache("acetone");
if (results) {
console.log(`Found ${results.length} products`);
}
// With custom limit
const results = await supplier.queryProductsWithCache("acetone", 10);
if (results) {
for (const builder of results) {
const product = await builder.build();
console.log(product.title, product.price);
}
}
protected async queryProductsWithCache(
query: string,
limit: number = this.limit,
): Promise<ProductBuilder<T>[] | void> {
// Check cache first (processed product data)
this.logger.debug(
'queryProductsWithCache: called for',
this.supplierName,
'query:',
query,
'limit:',
limit,
);
const key = this.cache.generateCacheKey(query);
const cached = await this.cache.getCachedQueryEntry(key);
this.logger.debug('queryProductsWithCache: cache hit:', !!cached, 'key:', key);
if (cached) {
const cachedLimit = cached.__cacheMetadata.limit;
const insufficientLimit = typeof cachedLimit === 'number' && cachedLimit < limit;
if (!insufficientLimit) {
this.logger.debug('Returning cached query results');
// Re-initialize product builders from cached processed data
return ProductBuilder.createFromCache<T>(this.baseURL, cached.data.slice(0, limit));
}
// Cached entry was built with a smaller limit than requested — drop it and re-query below.
this.logger.debug('Invalidating query cache due to insufficient limit', {
cachedLimit,
requestedLimit: limit,
});
await deleteSupplierQueryCacheEntry(key);
}
// If not in cache, perform the actual query. Run setup first so any
// subclass state it mutates (headers, localStorage, tokens, etc.) is
// in place before `queryProducts` reads it. Memoized, so this is cheap
// on repeat calls within the same supplier instance.
await this.ensureSetup();
const results = await this.queryProductsResolved(query, limit);
if (results) {
// Store processed results in cache (dumped/serialized form) and the limit used
await this.cache.cacheQueryResults(
query,
results.map((b) => b.dump()),
limit,
);
}
return results;
}
ProtectedqueryResolves the supplier's queryProducts for the current search, applying the
keyword-only advanced-search fallback when needed. For a plain query, or for
a supplier that handles boolean queries natively
(supportsNativeAdvancedSearch), this is a single queryProducts
call. For an advanced query on a keyword-only backend, it issues one search
per derived OR-group term (deriveFallbackTerms) and unions the
results, deduping by product URL/ID. Each queryProducts batch already
enforces the full boolean predicate via fuzzyFilterAst, so the union
is the set of products matching the whole query. Honors
httpRequestHardLimit; the fetches run inside execute()'s existing
maxAllowableSearchTimeSec race so an aborted controller cancels them.
The raw search query.
The per-supplier result limit.
The (possibly unioned) product builders, or void.
protected async queryProductsResolved(
query: string,
limit: number,
): Promise<ProductBuilder<T>[] | void> {
const parsed = this.getAst();
if (!parsed.isAdvanced || this.supportsNativeAdvancedSearch) {
return this.queryProducts(query, limit);
}
const terms = this.deriveFallbackTerms();
if (terms.length <= 1) {
// Nothing to union (single positive term, or a purely negative query):
// run the raw query once and let fuzzyFilterAst enforce the predicate.
return this.queryProducts(terms[0] ?? query, limit);
}
const seen = new Set<string>();
const union: ProductBuilder<T>[] = [];
for (const term of terms) {
if (this.requestCount >= this.httpRequestHardLimit) {
this.logger.warn('queryProductsResolved: httpRequestHardLimit reached, stopping fallback', {
term,
requestCount: this.requestCount,
});
break;
}
const batch = await this.queryProducts(term, limit);
if (!batch) {
continue;
}
for (const builder of batch) {
const key = String(builder.get('url') ?? builder.get('id') ?? builder.get('title') ?? '');
if (key === '' || seen.has(key)) {
continue;
}
seen.add(key);
union.push(builder);
}
}
return union.length > 0 ? union : undefined;
}
Executes the supplier's search query and returns the results. This method will execute all results concurrently (to the limits set in the supplier class), and resolve to an array of product objects.
Promise resolving to an array of products
This method is used to execute the supplier's search query and return the results.
public async *execute(): AsyncGenerator<T, void, undefined> {
// setup() is not called eagerly here — it's run lazily from the
// phase-boundary gates inside `queryProductsWithCache` and
// `getProductData` / `getProductDataWithCache`. A fully cached search
// never reaches those gates, so setup's token/cookie/permission
// requests are skipped entirely.
// Snapshot the user's ignore list once per search. Any product whose
// exclusion key matches an entry here is dropped before the detail phase
// runs (see the filter after queryProductsWithCache below).
this.excludedProductKeys = await loadExcludedProductKeys();
// Over-fetch by the number of previously-ignored products belonging to
// this supplier so that, in the worst case where every ignored product
// appears in the top of the query result set, we still end up with
// `this.limit` survivors after filtering. The queryProductsWithCache
// cache invalidates itself when the requested limit exceeds the cached
// limit, so this is safe.
const excludedForSupplier = await countExcludedProductsForSupplier(this.supplierName);
const fetchLimit = this.limit + excludedForSupplier;
incrementSearchQueryCount(this.supplierName);
// Optional per-supplier search-time budget; see armSearchTimeout. When it elapses the
// race below wins via SEARCH_TIMEOUT and flushes any not-yet-yielded products.
const SEARCH_TIMEOUT = Symbol('searchTimeout');
const { promise: timeoutPromise, handle: timeoutHandle } =
this.armSearchTimeout(SEARCH_TIMEOUT);
try {
const results = await this.queryProductsWithCache(this.query, fetchLimit);
if (!results || results.length === 0) {
this.logger.log(`No query results found`);
return;
}
// Drop any products the user has ignored, then slice back down to the
// user-visible limit. Uses the same dual-read (identity + legacy URL) as
// getProductData/partitionForBatch so the check is consistent wherever it
// fires first.
const survivors: ProductBuilder<T>[] = [];
for (const builder of results) {
if (survivors.length >= this.limit) break;
if (this.isExcluded(builder)) {
this.logger.debug('Skipping excluded product (pre-detail)', {
url: builder.get('url'),
});
continue;
}
survivors.push(builder);
}
this.products = survivors;
const queue = new Queue(this.maxConcurrentRequests, this.minConcurrentCycle);
// Each task fetches a product's detail data and finishes it, tagged with its index so the
// yield loop can track which products are still outstanding when the budget elapses.
const pending = new Map<number, Promise<{ index: number; finished: Maybe<T> }>>();
this.products.forEach((product, index) => {
pending.set(
index,
queue.run(async () => {
// If the budget already elapsed, skip the (now-aborted) detail fetch — the timeout
// handler below emits this product's basic data directly.
if (this.controller.signal.aborted) {
return { index, finished: undefined };
}
try {
const builder = await this.getProductData(product);
const finished = builder ? await this.finishProduct(builder) : undefined;
return { index, finished };
} catch (e: unknown) {
this.logger.error('Error processing product', { error: e, product });
incrementParseError(this.supplierName);
return { index, finished: undefined };
}
}),
);
});
// As each task resolves, yield the product. Race against the optional search-time budget;
// when it elapses, emit every not-yet-yielded product with its basic (query-phase) data so
// the rows still show — the enrichment for those simply wasn't cached, so a later search
// (served from the query cache) re-fetches their detail data.
const yielded = new Set<number>();
while (pending.size > 0) {
const result = await Promise.race(
timeoutPromise ? [...pending.values(), timeoutPromise] : [...pending.values()],
);
if (result === SEARCH_TIMEOUT) {
for (let index = 0; index < this.products.length; index++) {
if (yielded.has(index)) continue;
// The builder is enriched in place, so this carries whatever detail data was set
// before the abort, falling back to the basic query-phase fields otherwise.
const finished = await this.finishProduct(this.products[index]);
if (finished) {
yield finished;
}
}
break;
}
pending.delete(result.index);
yielded.add(result.index);
if (result.finished) {
yield result.finished;
}
}
} finally {
if (timeoutHandle !== undefined) {
clearTimeout(timeoutHandle);
}
}
}
ProtectedfinishFinalizes a partial product by adding computed properties and validating the result. This method:
The ProductBuilder instance containing the partial product to finalize
Promise resolving to a complete Product object or void if validation fails
// Example with a valid partial product
const builder = new ProductBuilder<Product>(this.baseURL);
builder
.setBasicInfo("Sodium Chloride", "/products/nacl", "ChemSupplier")
.setPricing(29.99, "USD", "$")
.setQuantity(500, "g");
const finishedProduct = await this.finishProduct(builder);
if (finishedProduct) {
console.log("Finalized product:", {
title: finishedProduct.title,
price: finishedProduct.price,
quantity: finishedProduct.quantity,
uom: finishedProduct.uom,
usdPrice: finishedProduct.usdPrice,
baseQuantity: finishedProduct.baseQuantity
});
}
// Example with an invalid partial product
const invalidBuilder = new ProductBuilder<Product>(this.baseURL);
invalidBuilder.setBasicInfo("Sodium Chloride", "/products/nacl", "ChemSupplier");
// Missing required fields
const invalidProduct = await this.finishProduct(invalidBuilder);
if (!invalidProduct) {
console.log("Failed to finalize product - missing required fields");
}
protected async finishProduct(product: ProductBuilder<T>): Promise<Maybe<T>> {
if (!isMinimalProduct(product.dump())) {
this.logger.warn('Unable to finish product - Minimum data not set', { product });
return;
}
// Dev guard: every builder should carry a stamped `cacheKey` (set at parse
// time via `setCacheKey(getUniqueProductKey(item))`). A missing key means a
// supplier implemented `getUniqueProductKey` but forgot to stamp — it would
// silently disable per-product caching and precise exclusion for that
// product. Surface it loudly in dev/tests.
if (IS_DEV_BUILD && product.get('cacheKey') == null) {
this.logger.error('finishProduct| product is missing a stamped cacheKey', {
supplier: this.supplierName,
url: product.get('url'),
});
}
// Set the country and shipping scope of the supplier
// have different restrictions on different products or countries.
product.setSupplierCountry(this.country);
product.setSupplierShipping(this.shipping);
if (this.paymentMethods.length > 0) {
product.setSupplierPaymentMethods(this.paymentMethods);
}
// Marketplace storefronts. The "*only" methods point users away from the supplier's own
// (restricted) site, so a missing store URL is a misconfiguration — surface it loudly in dev.
// The plain "ebay"/"amazon" methods instead drive an informational "more products there"
// notice; the store URL is optional, so stamp it only when present and don't warn.
if (this.paymentMethods.includes('ebayonly')) {
if (IS_DEV_BUILD && !this.ebayStoreURL) {
this.logger.error("finishProduct| supplier declares 'ebayonly' but sets no ebayStoreURL", {
supplier: this.supplierName,
});
}
product.setSupplierEbayStoreURL(this.ebayStoreURL);
} else if (this.paymentMethods.includes('ebay') && this.ebayStoreURL) {
product.setSupplierEbayStoreURL(this.ebayStoreURL);
}
if (this.paymentMethods.includes('amazononly')) {
if (IS_DEV_BUILD && !this.amazonStoreURL) {
this.logger.error(
"finishProduct| supplier declares 'amazononly' but sets no amazonStoreURL",
{ supplier: this.supplierName },
);
}
product.setSupplierAmazonStoreURL(this.amazonStoreURL);
} else if (this.paymentMethods.includes('amazon') && this.amazonStoreURL) {
product.setSupplierAmazonStoreURL(this.amazonStoreURL);
}
const built = await product.build();
return built;
}
ProtectedproductThe stable per-product cache/exclusion key for a builder: the identity
stamped on it at parse time (getUniqueProductKey →
setCacheKey), hashed with the supplier name via
getProductIdentityKey. Returns undefined when no identity was
stamped (so callers can skip the identity cache/exclusion path).
The product builder
The identity cache key, or undefined when unstamped
const key = this.productIdentityKey(builder); // md5({key, supplier}) or undefined
protected productIdentityKey(product: ProductBuilder<T>): string | undefined {
const identity = product.get('cacheKey');
if (typeof identity === 'string' && identity.length > 0) {
return this.cache.getProductIdentityCacheKey(identity);
}
return undefined;
}
ProtectedisWhether a product is on the user's ignore list, matched by its identity key (productIdentityKey) — the same key the "Ignore Product" action writes.
The product builder to check
true if the product matches an ignore-list entry
if (this.isExcluded(builder)) continue; // skip ignored product
protected isExcluded(product: ProductBuilder<T>): boolean {
const identityKey = this.productIdentityKey(product);
return identityKey !== undefined && this.excludedProductKeys.has(identityKey);
}
ProtectedpartitionPartitions query-phase builders for a batch supplier (one that enriches
details up front rather than per-product in getProductData). Runs a
three-way split: ignored products are dropped from both results;
cache hits (found in the product-detail cache by their stamped
identity) are hydrated in place via setData and kept in survivors but
excluded from misses; everything else is a miss, kept in both. The
caller enriches only misses, then caches them, and returns survivors.
Skips the cache lookup entirely when skipProductDetailCache is true (pure-search suppliers), so those still drop ignored products but treat every survivor as needing no enrichment.
The query-phase builders (each already setCacheKey-stamped)
{ survivors, misses } — see above
const { survivors, misses } = await this.partitionForBatch(builders);
await this.enrichVariants(misses);
await this.cacheProductBuilders(misses);
return survivors;
protected async partitionForBatch(
products: ProductBuilder<T>[],
): Promise<{ survivors: ProductBuilder<T>[]; misses: ProductBuilder<T>[] }> {
const survivors: ProductBuilder<T>[] = [];
const misses: ProductBuilder<T>[] = [];
for (const product of products) {
if (this.isExcluded(product)) {
continue;
}
survivors.push(product);
if (this.skipProductDetailCache) {
continue;
}
const key = this.productIdentityKey(product);
const cached = key ? await this.cache.getCachedProductData(key) : undefined;
if (isCachedProductData<T>(cached)) {
product.setData(cached);
} else {
misses.push(product);
}
}
return { survivors, misses };
}
ProtectedcacheWrites each enriched builder to the product-detail cache under its stamped
identity key. No-ops for suppliers with skipProductDetailCache true,
for aborted searches, and for builders shouldCacheProductData
rejects (e.g. a fetch that hit a noCacheStatusCode). Used by batch
suppliers after enriching their misses.
The enriched builders to persist
A promise that resolves once all writes complete
await this.cacheProductBuilders(misses);
protected async cacheProductBuilders(products: ProductBuilder<T>[]): Promise<void> {
if (this.skipProductDetailCache || this.controller.signal.aborted) {
return;
}
await Promise.all(
products.map(async (product) => {
const key = this.productIdentityKey(product);
if (key && this.shouldCacheProductData(product)) {
await this.cache.cacheProductData(key, product.dump());
}
}),
);
}
ProtectedhrefTakes in either a relative or absolute URL and returns an absolute URL. This is useful for when you aren't sure if the link (retrieved from parsed text, a setting, an element, an anchor value, etc) is absolute or not. Using relative links will result in http://chrome-extension://... being added to the link.
URL object or string
Optionalparams: Maybe<RequestParams>The parameters to add to the URL.
Optionalhost: stringThe host to use for overrides (eg: needing to call a different host for an API)
absolute URL
this.href('/some/path')
// https://supplier_base_url.com/some/path
this.href('https://supplier_base_url.com/some/path', null, 'another_host.com')
// https://another_host.com/some/path
this.href('/some/path', { a: 'b', c: 'd' }, 'another_host.com')
// http://another_host.com/some/path?a=b&c=d
this.href('https://supplier_base_url.com/some/path')
// https://supplier_base_url.com/some/path
this.href(new URL('https://supplier_base_url.com/some/path'))
// https://supplier_base_url.com/some/path
this.href('/some/path', { a: 'b', c: 'd' })
// https://supplier_base_url.com/some/path?a=b&c=d
this.href('https://supplier_base_url.com/some/path', new URLSearchParams({ a: 'b', c: 'd' }))
// https://supplier_base_url.com/some/path?a=b&c=d
protected href(path: string | URL, params?: Maybe<RequestParams>, host?: string): string {
const href = new URL(path, this.baseURL);
if (host) {
href.host = host;
}
if (params && Object.keys(params).length > 0) {
href.search = this.buildSearchParams(params).toString();
}
return String(href);
}
ProtectedbuildRecursively serializes an object into URLSearchParams, encoding nested
objects with bracket notation (parent[child]=value). Fixes the case that
new URLSearchParams(obj) mishandles — it stringifies a nested object to
"[object Object]" — so a params object with nested filters serializes
correctly. Primitive values (and arrays, which serialize comma-joined like
URLSearchParams does) are appended directly.
The params object to serialize
The URLSearchParams to append into (defaults to a fresh instance)
The key prefix used while recursing into nested objects
The populated URLSearchParams
this.buildSearchParams({ q: "acid", filter: { size: "500g" } }).toString();
// "q=acid&filter%5Bsize%5D=500g" (i.e. filter[size]=500g)
protected buildSearchParams(
obj: Record<string, unknown>,
params: URLSearchParams = new URLSearchParams(),
prefix: string = '',
): URLSearchParams {
for (const [key, value] of Object.entries(obj)) {
// Bracket-nest the key as depth grows: parent[child][grandchild]...
const formKey = prefix ? `${prefix}[${key}]` : key;
if (isPopulatedObject(value)) {
this.buildSearchParams(value, params, formKey);
} else {
params.append(formKey, String(value));
}
}
return params;
}
ProtectedgetRetrieves product data with caching support. Similar to getProductData but allows for additional parameters to be included in the cache key.
The ProductBuilder instance to get data for
The function to use for fetching product data
Promise resolving to the updated ProductBuilder or void if fetch fails
const builder = new ProductBuilder<Product>(this.baseURL);
builder.setBasicInfo("Acetone", "/products/acetone", "ChemSupplier");
const updatedBuilder = await supplier.getProductDataWithCache(
builder,
async (b) => {
// Custom fetching logic
return b;
},
);
protected async getProductDataWithCache(
product: ProductBuilder<T>,
fetcher: (builder: ProductBuilder<T>) => Promise<ProductBuilder<T> | void>,
): Promise<ProductBuilder<T> | void> {
const url = product.get('url');
if (typeof url !== 'string') {
this.logger.error('[SupplierBase > getProductDataWithCache] Invalid URL in product:', {
url,
});
return undefined;
}
// Skip products the user has ignored (matched by identity key).
if (this.isExcluded(product)) {
this.logger.debug('[SupplierBase > getProductDataWithCache] Skipping excluded product', {
url,
supplierName: this.supplierName,
});
return undefined;
}
// Key by the supplier's stable identity, stamped on the builder at parse
// time. Absent only if a supplier failed to stamp one, in which case this
// product simply isn't cached.
const cacheKey = this.productIdentityKey(product);
this.logger.debug(
'[SupplierBase > getProductDataWithCache] Product detail cache key:',
cacheKey,
{
url,
},
);
try {
if (!this.skipProductDetailCache && cacheKey) {
const cachedData = await this.cache.getCachedProductData(cacheKey);
if (isCachedProductData<T>(cachedData)) {
product.setData(cachedData);
return product;
}
}
// Cache miss (or caching skipped): run setup (memoized) so any state
// subclasses rely on is ready before the fetcher reads it, then fetch.
await this.ensureSetup();
let resultBuilder: ProductBuilder<T> | void = undefined;
try {
resultBuilder = await fetcher(product);
} catch (err: unknown) {
this.logger.error(
'[SupplierBase > getProductDataWithCache] Error in product detail fetcher:',
err,
);
incrementParseError(this.supplierName);
return undefined;
}
if (resultBuilder) {
incrementProductCount(this.supplierName);
// Skip caching when the search was aborted (e.g. maxAllowableSearchTimeSec) — the enrichment
// fetch was cancelled, so the data is incomplete and a later search should retry it.
if (
cacheKey &&
!this.skipProductDetailCache &&
!this.controller.signal.aborted &&
this.shouldCacheProductData(resultBuilder)
) {
await this.cache.cacheProductData(cacheKey, resultBuilder.dump());
}
}
return resultBuilder;
} catch (outerErr: unknown) {
this.logger.error(
'[SupplierBase > getProductDataWithCache] Error in getProductDataWithCache:',
outerErr,
);
incrementParseError(this.supplierName);
return undefined;
}
}
ProtectedproductThe key under which a product's detail-fetch failures are recorded/looked up: its permalink if set, otherwise its processing URL. Subclasses that fetch supplemental data should record failures under this same key so shouldCacheProductData can match them.
The product builder
The fetch key, or undefined when neither permalink nor url is a string
this.productFetchKey(builder); // "https://www.aladdinsci.com/us_en/x.html"
protected productFetchKey(product: ProductBuilder<T>): string | undefined {
const key = product.get('permalink') ?? product.get('url');
return typeof key === 'string' ? key : undefined;
}
ProtectedrecordRecords the HTTP status of a failed product-detail fetch so shouldCacheProductData can skip caching it. Non-HTTP failures (no status) are not recorded.
The product fetch key (see productFetchKey)
OptionalhttpStatus: numberThe HTTP status of the failure, if it was an HttpError
void
this.recordFetchFailure(permalink, 429);
protected recordFetchFailure(key: string, httpStatus?: number): void {
if (typeof httpStatus === 'number') {
this.failedFetchStatuses.set(key, httpStatus);
}
}
ProtectedshouldDecides whether a freshly-fetched product's detail data should be written to the cache. Skips caching when the product's last detail fetch failed with a noCacheStatusCodes status (default 429), so a later search retries it instead of serving the incomplete cached entry. The product is still listed regardless.
The product builder about to be cached
True to cache the product data, false to skip caching it
this.shouldCacheProductData(builder); // false when the detail fetch hit a 429
protected shouldCacheProductData(product: ProductBuilder<T>): boolean {
const key = this.productFetchKey(product);
if (key === undefined) {
return true;
}
const failedStatus = this.failedFetchStatuses.get(key);
return failedStatus === undefined || !this.noCacheStatusCodes.includes(failedStatus);
}
ProtectedgroupGroups variants of a product by their title
Array of product listings from search results
Array of product listings with grouped variants
Create a generic method for this, the same method is used in Synthetika and could be of use with LoudWolf.
const results = await this.queryProducts("sodium chloride");
const grouped = this.groupVariants(results);
// grouped is an array of product listings with grouped variants
protected groupVariants<R>(data: R[]): R[] {
const variants: GroupedItem<R>[] = data
.map((item) => {
const title = this.titleSelector(item);
if (!title) {
this.logger.error('No title found in product:', { item });
return undefined;
}
const groupId = stripQuantityFromString(title.replace(/(?<=\d{1,3})\s(?=\d{3})/g, ''));
const groupIdWithoutSpaces = groupId.replace(/[\s-]/g, '');
return { ...item, groupId: groupIdWithoutSpaces };
})
.filter((item): item is GroupedItem<R> => item !== undefined);
const products = Object.groupBy(variants, (item) => item.groupId);
return Object.values(products)
.filter((product): product is GroupedItem<R>[] => product !== undefined)
.map((product) => {
const main = product.splice(0, 1)[0];
// eslint-disable-next-line @typescript-eslint/no-unused-vars
const { groupId, ...newObject } = main;
// R is unconstrained, so TS widens newObject.variants to
// R["variants"] & (R[] | undefined), which it cannot prove product
// (GroupedItem<R>[]) satisfies; assert to the destructured property type.
newObject.variants = product as typeof newObject.variants;
return newObject;
})
.filter((item): item is GroupedItem<R> => item !== undefined);
}
ProtectedbackgroundRuns an HTTP request from the extension's background service worker instead of this
(page) context, sidestepping the CORS restrictions that apply to extension pages.
The target host must be granted in the manifest host_permissions. Returns a real
Response (text/JSON bodies only). Independent of fetch — it does not
share the request counter, hard limit, or WAF retry logic.
The absolute URL to request.
Optionalinit: BackgroundFetchInitOptional serializable request options (method, headers, body, etc.).
A Response reconstructed from the worker's reply.
// Inside a supplier method, e.g. scraping an asset blocked by page CORS:
const homepage = await this.backgroundFetch("https://chemsavers.com/");
const html = await homepage.text();
const search = await this.backgroundFetch("https://api.example.com/search", {
method: "POST",
headers: { "content-type": "application/json" },
body: JSON.stringify({ q: "acid" }),
});
const results = search.ok ? await search.json() : undefined;
protected async backgroundFetch(url: string, init?: BackgroundFetchInit): Promise<Response> {
this.logger.debug(`Background fetching: ${url}`);
return backgroundFetch(url, init);
}
ProtectedfetchInternal fetch method with request counting and decorator. Tracks request count and enforces hard limits on HTTP requests.
Arguments to pass to fetchDecorator (usually a Request or URL and options)
The response from the fetchDecorator
Error if request count exceeds hard limit
// Example usage inside a subclass:
const response = await this.fetch(new Request('https://example.com'));
if (response.ok) {
const data = await response.json();
console.log(data);
}
// With custom request options
const response = await this.fetch(
new Request('https://example.com', {
headers: { 'Accept': 'application/json' }
})
);
protected async fetch(
...args: Parameters<typeof fetchDecorator>
): Promise<FetchDecoratorResponse> {
const [input] = args;
this.logger.debug(`Fetching: ${input}`);
// One initial attempt plus up to `challengeRetryLimit` retries. A 403 from
// a WAF cookie handshake plants a cookie on the first hit (stored because
// credentials:"include"); the retry sends it back and usually passes.
const maxAttempts = 1 + Math.max(0, this.challengeRetryLimit);
for (let attempt = 1; attempt <= maxAttempts; attempt++) {
// Each attempt is a real network request, so it counts toward the hard
// limit. For non-retrying suppliers (maxAttempts === 1) this is
// identical to the previous single increment.
this.requestCount++;
if (this.requestCount > this.httpRequestHardLimit) {
this.logger.warn('Request count exceeded hard limit', { requestCount: this.requestCount });
incrementFailure(this.supplierName);
throw new Error('Request count exceeded hard limit');
}
try {
const response = await fetchDecorator(...args);
this.logger.debug(`Response Status: ${response.status}`);
this.logger.debug('response hash:', response.requestHash);
if (typeof response.data === 'string' && response.data?.length === 0) {
throw new EmptyResponseError(`Invalid response: ${response.data}`);
}
incrementSuccess(this.supplierName);
return response;
} catch (error: unknown) {
if (this.shouldRetryChallenge(error) && attempt < maxAttempts) {
this.logger.warn('Retrying after 403 (WAF cookie handshake)', {
attempt,
maxAttempts,
input,
});
await sleep(this.challengeRetryDelayMs);
continue;
}
incrementFailure(this.supplierName);
throw error;
}
}
// Unreachable: the loop always returns on success or throws on the final
// failed attempt. Present only to satisfy the return-type checker.
throw new Error('fetch: exhausted retries without resolving');
}
ProtectedgetDerives the stable unique key for a ScienceLab search item: its URL slug, which survives the query→detail transition and is unique per product.
The raw ScienceLabItem
The product slug
this.getUniqueProductKey(item); // "sodium-hexametaphosphate-anhydrous"
protected getUniqueProductKey(data: ScienceLabItem): string {
return data.slug;
}
ProtectedtitleExtracts the humanized product name from a ScienceLab search item. Used by the base fuzzy filter to score each item against the query.
The raw ScienceLabItem
The humanized product name, or undefined when absent
this.titleSelector(item); // "sodium hexametaphosphate anhydrous"
protected titleSelector(data: ScienceLabItem): Maybe<string> {
return data?.name;
}
ProtectedqueryQueries ScienceLab products. Fetches the full catalog from the XML sitemap,
fuzzy-filters the humanized names against the query, and returns basic
builders for the top limit matches (priced later in getProductData).
The search term to query products for
The maximum number of results to return
Basic product builders for the top matches, or void on failure
const results = await supplier.queryProducts("acetone", 5);
console.log(results?.length); // up to 5
protected async queryProducts(
query: string,
limit: number = this.limit,
): Promise<ProductBuilder<Product>[] | void> {
this.logger.log('queryProducts:', { query, limit });
const catalog = await this.fetchCatalog();
if (catalog.length === 0) {
this.logger.error('Empty ScienceLab catalog', { query });
return;
}
// Score every catalog entry directly and rank by score. The base
// `fuzzyFilter`/`extract` path can't be used here: on a list this large its
// fuzzball `extract` call misassigns scores to the wrong items (e.g. an
// unrelated title reported at 95). `fuzzyScoreAst` scores one title at a
// time against the parsed query — honoring the query AST, the user's scorer
// override, and the match cutoff (returning null to drop) — so it stays
// correct at scale.
const matches: ScienceLabItem[] = [];
for (const item of catalog) {
const score = this.fuzzyScoreAst(item.name);
if (score === null) {
continue;
}
item.matchPercentage = score;
matches.push(item);
}
matches.sort((a, b) => (b.matchPercentage ?? 0) - (a.matchPercentage ?? 0));
// Scored per item via `fuzzyScoreAst`, so the base `fuzzyFilter` scorer
// comparison table never fires — log the ranked head instead for visibility.
this.logger.debug('queryProducts ranked:', {
query,
matched: matches.length,
top: matches.slice(0, 10).map((m) => ({ score: m.matchPercentage, name: m.name })),
});
return this.initProductBuilders(matches.slice(0, limit));
}
PrivatefetchFetches the whole ScienceLab catalog from the product XML sitemap. Walks the paginated sitemap until a page adds no new products (covering both an empty page and a page that repeats the previous one), deduping by slug.
Every catalog item, with a humanized name
const catalog = await this.fetchCatalog();
console.log(catalog[0]); // { url, slug, name }
private async fetchCatalog(): Promise<ScienceLabItem[]> {
const items: ScienceLabItem[] = [];
const seen = new Set<string>();
for (let page = 1; page <= this.maxSitemapPages; page++) {
const xml = await this.fetchSitemapPage(page);
if (!xml) {
break;
}
const before = items.length;
for (const url of this.extractLocs(xml)) {
const slug = slugFromUrl(url);
if (!slug || seen.has(slug)) {
continue;
}
seen.add(slug);
items.push({ url, slug, name: humanizeSlug(slug) });
}
// No new products on this page — end of catalog (or a repeated page).
if (items.length === before) {
break;
}
}
return items;
}
PrivatefetchFetches one page of the product sitemap as raw XML. Uses httpGet rather
than httpGetHtml because the sitemap is served as XML, which the HTML
content-type guard may reject.
The 1-based sitemap page number
The page XML, or undefined when the request fails
const xml = await this.fetchSitemapPage(1);
private async fetchSitemapPage(page: number): Promise<Maybe<string>> {
const response = await this.httpGet({
path: '/xmlsitemap.php',
params: { type: 'products', page },
});
if (!isHttpResponse(response) || !response.ok) {
return undefined;
}
return await response.text();
}
PrivateextractExtracts every <loc> URL from a sitemap XML string.
The sitemap XML
The URLs, trimmed
this.extractLocs("<loc>https://sciencelab.com/acetone/</loc>"); // ["https://sciencelab.com/acetone/"]
private extractLocs(xml: string): string[] {
const matches = xml.match(/<loc>\s*([^<]+?)\s*<\/loc>/g) ?? [];
return matches.map((loc) => loc.replace(/<\/?loc>/g, '').trim());
}
ProtectedinitBuilds basic product builders (title, URL, cache key, match score) from
ranked catalog items. Pricing and specs are filled later by getProductData.
The ranked ScienceLabItem matches
One builder per item
const builders = this.initProductBuilders(matches);
protected initProductBuilders(items: ScienceLabItem[]): ProductBuilder<Product>[] {
return mapDefined(items, (item) =>
new ProductBuilder<Product>(this.baseURL)
.setBasicInfo(item.name, item.url, this.supplierName)
.setCacheKey(this.getUniqueProductKey(item))
.setMatchPercentage(item.matchPercentage),
);
}
ProtectedgetFetches and parses a ScienceLab product page: the ld+json Product block (real title, sku, image, price, availability), the spec table (CAS, molecular formula, molecular weight, special-considerations restriction), the grade and concentration (from the title), and each size variant (priced via the product-attributes endpoint).
The basic product builder from queryProducts
The enriched builder, or void when the page can't be fetched
const full = await supplier.getProductData(builder);
console.log(full?.dump().cas, full?.dump().variants?.length);
protected async getProductData(
product: ProductBuilder<Product>,
): Promise<ProductBuilder<Product> | void> {
return this.getProductDataWithCache(product, async (builder) => {
if (typeof builder === 'undefined') {
this.logger.error('No products to get data for', { builder });
return;
}
const html = await this.httpGetHtml({ path: builder.get('url') });
if (!html) {
this.logger.warn('No product response', { builder });
return;
}
const dom = createDOM(html);
// schema.org Product ld+json: real title, sku, image, price, availability.
const schema = SchemaOrgData.fromDocument(dom);
const productNode = schema.first('Product');
const offer = schema.all('Offer')[0];
const description = this.decodeDescription(productNode?.description);
if (productNode) {
// The real store title replaces the humanized slug guess.
builder.setTitle(productNode.name);
builder.setSku(productNode.sku);
builder.setID(productNode.mpn);
builder.setImage(
Array.isArray(productNode.image) ? productNode.image[0] : productNode.image,
);
if (description) {
builder.setDescription(description);
}
}
this.applyPricing(builder, offer, dom);
const availability =
offer && typeof offer.availability === 'string'
? offer.availability
: this.metaContent(dom, 'og:availability');
if (availability) {
builder.setAvailability(availability);
}
this.applyInfoFields(builder, dom, description);
const title = String(builder.get('title') ?? '');
builder.setGrade(parseGrade(title));
builder.setConcentration(parseConcentration(title));
// BigCommerce product id (needed for variant pricing) lives only in og:id.
const productId = this.metaContent(dom, 'og:id');
if (productId) {
const variants = await this.parseVariants(dom, productId);
if (variants.length > 0) {
builder.setVariants(variants);
this.anchorBaseVariant(builder, variants[0]);
}
}
return builder;
});
}
PrivateapplySets the base product price and currency from the ld+json offer (its
minPrice — the smallest size), falling back to the product:price:amount
meta tag.
The product builder to price
The ld+json Offer node, if present
The parsed product page
private applyPricing(
builder: ProductBuilder<Product>,
offer: SchemaNode | undefined,
dom: Document,
): void {
const price = this.readOfferPrice(offer) ?? this.readMetaPrice(dom);
if (price === undefined) {
return;
}
const currencyCode =
offer && typeof offer.priceCurrency === 'string' ? offer.priceCurrency : 'USD';
builder.setPrice(price);
builder.setCurrencyCode(currencyCode);
builder.setCurrencySymbol(CURRENCY_SYMBOL_MAP[currencyCode]);
}
PrivatereadReads the lowest numeric price from a ld+json Offer node, trying minPrice,
price, then lowPrice (values may be numbers or numeric strings).
The ld+json Offer node, if present
The price, or undefined when none is parseable
this.readOfferPrice({ minPrice: "108" }); // 108
private readOfferPrice(offer: SchemaNode | undefined): number | undefined {
if (!offer) {
return undefined;
}
for (const key of ['minPrice', 'price', 'lowPrice']) {
const value = offer[key];
const num =
typeof value === 'number' ? value : typeof value === 'string' ? Number(value) : NaN;
if (Number.isFinite(num)) {
return num;
}
}
return undefined;
}
PrivatereadReads the product:price:amount meta tag as a fallback base price.
The parsed product page
The price, or undefined when absent/unparseable
this.readMetaPrice(dom); // 108
private readMetaPrice(dom: Document): number | undefined {
const raw = this.metaContent(dom, 'product:price:amount', 'property');
if (raw === undefined) {
return undefined;
}
const num = Number(raw);
return Number.isFinite(num) ? num : undefined;
}
PrivateapplyApplies the product-page spec table to the builder: CAS number, molecular formula, molecular weight, and "Special Considerations" (kept as an informational restriction note).
The spec table is the only trustworthy source for CAS and formula. When the
molecular-formula row is absent, the ld+json description is a safe fallback
— findFormulaInText yields the formula for a single compound and nothing
for a solution. There is deliberately no CAS fallback: a solution's
description leads with its solvent (water, acetic acid), so scraping a CAS
from it would mislabel the product with the solvent's number.
"Special Considerations" is a generic hazmat/handling notice ScienceLab
shows on most products ("higher shipping charges … may apply", "Customer
Service will contact you"). It is captured as a note only — never
restrictedDelivery/buyerRestricted, which canUserBuy treats as
un-buyable and would hide every product.
The product builder to enrich
The parsed product page
The decoded ld+json description, if any
private applyInfoFields(
builder: ProductBuilder<Product>,
dom: Document,
description: string | undefined,
): void {
let formulaSet = false;
for (const row of Array.from(dom.querySelectorAll('dl.productView-info .line-item-details'))) {
const name = row.querySelector('.productView-info-name')?.textContent?.trim().toLowerCase();
const value = row.querySelector('.productView-info-value')?.textContent?.trim();
if (!name || !value) {
continue;
}
if (name.includes('cas')) {
builder.setCAS(value);
} else if (name.includes('molecular formula')) {
// Format digit runs into subscripts (Na6O18P6 -> Na₆O₁₈P₆).
builder.setFormula(formatFormula(value));
formulaSet = true;
} else if (name.includes('molecular weight')) {
builder.setMoleweight(value);
} else if (name.includes('special considerations')) {
builder.setPurchaseRestriction({ note: value });
}
}
if (!formulaSet && description) {
const fallback = findFormulaInText(description);
if (fallback) {
builder.setFormula(formatFormula(fallback));
}
}
}
PrivateanchorAnchors the base product's quantity to its smallest size variant, so the
product carries a real quantity/uom. The base price is normally the
smallest size's price already (the ld+json minPrice), so the variant price
is used only as a last resort when no base price could be read.
The product builder
The smallest (first) size variant
private anchorBaseVariant(builder: ProductBuilder<Product>, variant: Partial<Variant>): void {
if (typeof variant.quantity === 'number' && variant.uom !== undefined) {
builder.setQuantity(variant.quantity, variant.uom);
}
if (builder.get('price') === undefined && typeof variant.price === 'number') {
builder.setPrice(variant.price);
}
}
PrivatecollectCollects the size-option choices from a product page. ScienceLab renders these in two styles, and a product uses one or the other:
input.form-radio per size, id
attribute_rectangle__{attributeId}_{valueId}, label in the paired
<label for>'s .form-option-variant.<select name="attribute[{attributeId}]"> with one
<option value="{valueId}"> per size (the empty "Choose Options"
placeholder is skipped).The parsed product page
One { attributeId, valueId, label } per size option
this.collectVariantOptions(dom);
// [{ attributeId: "1675", valueId: "4125", label: "500g" }, ...]
private collectVariantOptions(
dom: Document,
): { attributeId: string; valueId: string; label: string }[] {
const container = dom.querySelector('div.productView-options');
if (!container) {
return [];
}
const options: { attributeId: string; valueId: string; label: string }[] = [];
for (const radio of Array.from(container.querySelectorAll('input.form-radio'))) {
const id = radio.getAttribute('id') ?? '';
const idMatch = id.match(/__(\d+)_(\d+)$/);
const attributeId =
idMatch?.[1] ?? radio.getAttribute('name')?.match(/attribute\[(\d+)\]/)?.[1];
const valueId = idMatch?.[2] ?? radio.getAttribute('value')?.trim();
const label = id
? dom.querySelector(`label[for="${id}"] .form-option-variant`)?.textContent?.trim()
: undefined;
if (attributeId && valueId && label) {
options.push({ attributeId, valueId, label });
}
}
for (const select of Array.from(container.querySelectorAll('select'))) {
const attributeId = select.getAttribute('name')?.match(/attribute\[(\d+)\]/)?.[1];
if (!attributeId) {
continue;
}
for (const option of Array.from(select.querySelectorAll('option'))) {
const valueId = option.getAttribute('value')?.trim();
const label = option.textContent?.trim();
if (valueId && label) {
options.push({ attributeId, valueId, label });
}
}
}
return options;
}
PrivateparseParses a product's size variants (radio or dropdown, via collectVariantOptions) and prices each through the BigCommerce product-attributes endpoint.
The parsed product page
The BigCommerce product id (from the og:id meta)
The parsed variants (in the page's ascending-size order)
const variants = await this.parseVariants(dom, "1675");
// [{ id: "4125", title: "500g", quantity: 500, uom: "g", price: 108 }, ...]
private async parseVariants(dom: Document, productId: string): Promise<Partial<Variant>[]> {
const options = this.collectVariantOptions(dom);
if (options.length === 0) {
return [];
}
return Promise.all(
options.map(async ({ attributeId, valueId, label }): Promise<Partial<Variant>> => {
const qty = parseQuantity(label);
const price = await this.fetchVariantPrice(productId, attributeId, valueId);
return {
id: valueId,
title: label,
quantity: qty?.quantity,
uom: qty?.uom,
price,
};
}),
);
}
PrivatefetchPrices one size variant via the BigCommerce product-attributes AJAX endpoint
(POST remote/v1/product-attributes/{productId}), reading the selected
combination's data.price.without_tax.value.
The BigCommerce product id
The variant attribute group id
The attribute value id for the size to price
The variant price, or undefined when the request fails
await this.fetchVariantPrice("1675", "1675", "4126"); // 342
private async fetchVariantPrice(
productId: string,
attributeId: string,
valueId: string,
): Promise<number | undefined> {
const body = new URLSearchParams();
body.set('action', 'add');
body.set(`attribute[${attributeId}]`, valueId);
body.set('product_id', productId);
body.append('qty[]', '1');
try {
const response = await this.httpPostJson({
path: `/remote/v1/product-attributes/${productId}`,
body: body.toString(),
headers: {
'Content-Type': 'application/x-www-form-urlencoded; charset=UTF-8',
'X-Requested-With': 'stencil-utils',
},
});
if (!isScienceLabAttributeResponse(response)) {
this.logger.warn('Invalid variant response', { productId, attributeId, valueId });
return undefined;
}
const value = response.data?.price?.without_tax?.value;
return typeof value === 'number' ? value : undefined;
} catch (error) {
this.logger.warn('Variant price fetch failed', { productId, attributeId, valueId, error });
return undefined;
}
}
PrivatemetaReads a meta tag's content by its property (default) or name
attribute.
The parsed product page
The meta property/name value, e.g. "og:id"
Which attribute holds key ("property" or "name")
The trimmed content, or undefined when the tag is absent/empty
this.metaContent(dom, "og:id"); // "1675"
private metaContent(
dom: Document,
key: string,
attr: 'property' | 'name' = 'property',
): string | undefined {
const content = dom.querySelector(`meta[${attr}="${key}"]`)?.getAttribute('content')?.trim();
return content ? content : undefined;
}
PrivatedecodeDecodes the ld+json Product description, which ScienceLab percent-encodes.
The raw description value from the ld+json node
The decoded description, or undefined when not a string
this.decodeDescription("Sodium%20Hexametaphosphate"); // "Sodium Hexametaphosphate"
private decodeDescription(raw: unknown): string | undefined {
if (typeof raw !== 'string' || raw.length === 0) {
return undefined;
}
try {
return decodeURIComponent(raw);
} catch {
return raw;
}
}
Color used to visually tag this supplier's log output (and available for
charts/UI). Defaults to a stable palette color derived from the class name
via getSupplierColor, so no supplier has to set one. Override by
assigning a hex string in a subclass constructor (also call
this.logger.setColor(this.color) there to recolor the already-built logger).
Protected OptionalfuzzRuntime override resolved from userSettings.fuzzScorerOverride. When
set, fuzzyFilter uses this instead of this.fuzzScorer. Undefined
(the default) means "use whatever the supplier class picked". Mutated
by setFuzzScorerOverride so it can't be readonly.
Protected OptionalparsedParsed advanced-search query, set per-instance by SupplierFactory from the
user's input. When absent (e.g. a directly-constructed test supplier),
getAst lazily parses this.query instead.
Protected OptionalresolvedSMILES/structure query terms resolved to their chemical identifiers, keyed by
the raw search term. Resolved once per search by SupplierFactory (network-
bound, via NCI Cactus/PubChem) and shared with every supplier so none of them
re-resolve. Undefined when the query has no structure terms. Only suppliers
that filter by structure (currently Ambeed) read this; others ignore it.
Protected Static ReadonlysupportsWhether this supplier can search its own site by a CAS number / molecular formula / SMILES directly. Default false: the supplier matches product names, not identifiers, so when the query is one of these identifier types (see detectTermType) the base swaps it for the chemical name the factory resolved (see effectiveQuery). Set the relevant flag true on a supplier whose search natively accepts that identifier (e.g. Ambeed), so it receives the raw identifier unchanged.
Protected Static ReadonlysupportsSee supportsCAS.
Protected Static ReadonlysupportsSee supportsCAS.
ProtectedfuzzyRuntime flag resolved from userSettings.fuzzyFilteringDisabled. When true,
fuzzyFilterAst skips fuzzball scoring: plain queries return raw supplier
results and advanced queries are filtered only by the boolean predicate via
case-insensitive substring matching.
Protected ReadonlyfuzzyFuzzy strategy. When true (the default), fuzzyFilter/fuzzyFilterAst
rank candidates by fuzz score and keep them all (in score order) instead of dropping
anything below minMatchPercentage; the base search pipeline then caps the list
to limit, so the highest-scoring matches survive. This avoids dropping clear
matches whose ratio-style score falls under the cutoff purely because the title dwarfs
the query. Set to false on a supplier to restore the hard-cutoff behavior.
Protected ReadonlymaxMaximum number of backend search requests the keyword-only fallback issues.
Protected ReadonlyskipOpt-out flag for the per-product detail cache. Left false (the default),
every supplier caches its per-product detail data — the safe default, since
forgetting to set this just yields harmless redundant caching, never a
silent cache regression.
Set true only on a supplier that resolves every field in the initial
search (a passthrough getProductData with no per-product fetch): for those
the per-product cache saves nothing, so
getProductData/getProductDataWithCache/partitionForBatch
skip the product-detail cache read+write. The query cache still serves
repeat searches, and getUniqueProductKey is still used for
exclusions. Mark the concrete pure-search supplier (not a shared base
class), so a base's fetching subclass keeps caching by default.
Optional ReadonlyebayThe supplier's eBay storefront. Subclasses that list "ebayonly" in
SupplierBase.paymentMethods must override this; see
src/suppliers/__tests__/storeOnlyPaymentMethods.test.ts, which enforces the pairing that
TypeScript can't express.
Optional ReadonlyamazonThe supplier's Amazon storefront. Required alongside "amazononly", exactly as
ebayStoreURL is for "ebayonly".
Protected Static Optional ReadonlyshipsThe countries to which the supplier ships.
public readonly shipsTo: CountryCode[] = ["US", "CN", "NL"];
protected static readonly shipsTo?: CountryCode[];
Protected Static Optional ReadonlyapiOptional external API hostname used by some suppliers (e.g., Typesense,
Searchanise). When set, automatically included in requiredHosts for
permission checks.
ProtectedqueryString to query for (product name, CAS, etc.). The search term that will be used to find products. Set during construction and used throughout the supplier's lifecycle.
ProtectedqueryIf the products first require a query of a search page that gets iterated over, those results are stored here. Acts as a cache for the initial search results before they are processed into full product objects.
ProtectedbaseThe base search parameters that are always included in search requests. These parameters are merged with any additional search parameters when making requests to the supplier's API.
class MySupplier extends SupplierBase<Product> {
constructor() {
super();
this.baseSearchParams = {
format: "json",
version: "2.0"
};
}
}
protected baseSearchParams: Record<string, string | number> = {};
ProtectedcontrollerThe AbortController instance used to manage and cancel ongoing requests. This allows for cancellation of in-flight requests when needed, such as when a new search is started or the supplier is disposed.
const controller = new AbortController();
const supplier = new MySupplier("acetone", 5, controller);
// Later, to cancel all pending requests:
controller.abort();
protected controller: AbortController;
ProtectedlimitThe maximum number of results to return for a search query. This is not a limit on HTTP requests, but rather the number of products that will be returned to the caller.
const supplier = new MySupplier("acetone", 5); // Limit to 5 results
for await (const product of supplier) {
// Will yield at most 5 products
}
protected limit: number;
ProtectedproductsThe products that are currently being built by the supplier. This array holds ProductBuilder instances that are in the process of being transformed into complete Product objects.
await supplier.queryProducts("acetone");
console.log(`Building ${supplier.products.length} products`);
for (const builder of supplier.products) {
const product = await builder.build();
console.log("Built product:", product.title);
}
protected products: ProductBuilder<T>[] = [];
ProtectedrequestCounter for HTTP requests made during the current query execution. This is used to track the number of requests and ensure we don't exceed the httpRequestHardLimit.
0
await supplier.queryProducts("acetone");
console.log(`Made ${supplier.requestCount} requests`);
if (supplier.requestCount >= supplier.httpRequestHardLimit) {
console.log("Reached request limit");
}
protected requestCount: number = 0;
ProtectedminMinimum number of milliseconds between two consecutive tasks
protected minConcurrentCycle: number = 100;
ProtectedmaxMaximum wall-clock time (in seconds) a single supplier's execute() search may run.
Once exceeded, any in-flight and pending product-detail requests are aborted and the search
stops yielding new products — only those already collected are returned. Measured from the
start of execute(), so it also bounds a slow initial query. Set to 0 (the default) to
disable the limit. Override per supplier for sources that are slow or rate-limit-prone.
0 (disabled)
protected maxAllowableSearchTimeSec: number = 0;
ProtectedheadersHTTP headers used as a basis for all requests to the supplier. These headers are merged with any request-specific headers when making HTTP requests.
class MySupplier extends SupplierBase<Product> {
constructor() {
super();
this.headers = {
"Accept": "application/json",
"User-Agent": "ChemPal/1.0"
};
}
}
protected headers: HeadersInit = {};
Protected ReadonlyrequiredCookies that must be written into the browser jar before any request runs
— e.g. a currency or session-preference cookie the backend reads. Seeded
once per instance by ensureSetup (before setup) via chrome.cookies,
since the Cookie request header is on the fetch-forbidden list and can't
be set through this.headers. Each entry's url defaults to baseURL.
Subclasses override this instead of hand-rolling a setup that calls
chrome.cookies.set directly.
[]
class MySupplier extends SupplierBase<Partial<Product>, Product> {
protected readonly requiredCookies: SupplierCookieSeed[] = [
{ name: "currency", value: "2" },
];
}
protected readonly requiredCookies: SupplierCookieSeed[] = [];
Protected ReadonlychallengeNumber of times fetch retries a request that comes back 403. Some
suppliers sit behind a WAF that 403s the first hit while planting a
session cookie (a "cookie handshake"); because every request now sets
credentials: "include", that cookie lands in the jar and the retry
carries it back, usually passing. We can't gate on the Set-Cookie
header (it's fetch-forbidden and invisible to JS), so this per-supplier
flag is the gate — 0 (the default) means never retry. Only enable it
for suppliers known to do this handshake.
0
protected readonly challengeRetryLimit: number = 0;
Protected ReadonlychallengeDelay in milliseconds between 403 challenge retries. Gives the WAF a
brief beat before re-requesting with the freshly-planted cookie.
300
protected readonly challengeRetryDelayMs: number = 300;
ProtectedloggerLogger for the supplier. Initialized in the constructor with the name of the inheriting class.
ProtectedproductDefault values for products. These will get overridden if they're found in the product data.
ProtectedcacheCache instance for this supplier.
Initialized after construction by initCache() (called from
SupplierFactory once supplierName is set). The ! assertion is safe
here because every code path that reads this.cache
(queryProductsWithCache, getProductData, getProductDataWithCache)
runs only after execute() is called on a factory-built instance, and the
factory always calls initCache() before handing the instance out.
ProtectednoHTTP status codes that, when hit while fetching a product's detail data, prevent that
product from being cached (see shouldCacheProductData). Mirrors
userSettings.noCacheStatusCodes; set by initCache. Defaults to [429].
Protected ReadonlyfailedMaps a product's fetch key (permalink, falling back to its processing URL) to the HTTP status of its last failed detail fetch. Populated by subclasses via recordFetchFailure and consulted by shouldCacheProductData. Per-search, since the factory builds a fresh supplier instance for each search.
ProtectedexcludedProduct-data cache keys the user has explicitly excluded via the "Ignore
Product" context menu action. Loaded once per execute() from
storage.local so membership checks are synchronous on the hot path
(see getProductData). Newly-ignored products take effect on the next
search, which matches the stated feature requirement.
Static ReadonlysupplierThe name of the supplier (used for display name, lists, etc).
Static ReadonlybaseThe base URL for the supplier.
Static ReadonlyshippingThe shipping scope of the supplier. Used to determine the shipping scope of the supplier.
Static ReadonlycountryThe country code of the supplier. Used to determine the currency and other country-specific information.
Static ReadonlypaymentThe payment methods accepted by the supplier. Used to determine the payment methods accepted by the supplier.
ProtectedhttpMaximum number of HTTP requests allowed per search query. This is a hard limit to prevent excessive requests to the supplier's API. If this limit is reached, the supplier will stop making new requests.
50
class MySupplier extends SupplierBase<Product> {
constructor() {
super();
this.httpRequestHardLimit = 100; // Allow more requests
}
}
protected httpRequestHardLimit: number = 150;
ProtectedmaxNumber of requests to process in parallel when fetching product details. This controls the batch size for concurrent requests to avoid overwhelming the supplier's API and the user's bandwidth.
10
class MySupplier extends SupplierBase<Product> {
constructor() {
super();
// Process 5 requests at a time
this.maxConcurrentRequests = 5;
}
}
protected maxConcurrentRequests: number = 5;
Protected ReadonlyfuzzFuzz scorer used by fuzzyFilter to score each candidate's title against
the query. Any function from fuzzball with the
(str1, str2, opts?) => number shape works. Subclasses override this when
a supplier's title format needs a different scorer (e.g. a catalog that
pads titles with boilerplate might prefer partial_ratio). Defaults to
WRatio.
Overridable at runtime from userSettings.fuzzScorerOverride — see
setFuzzScorerOverride and fuzzyFilter below. The user's Advanced
settings selection wins over this subclass default when set.
Protected ReadonlyminThe minimum match percentage for a product to be considered a match.
Protected ReadonlysupportsWhether queryProducts handles an advanced (boolean) query natively in a
single request — true for suppliers that translate the AST into a server-side
query (Wix, Shopify, Chemsavers/Typesense, LiMac/FreeFind). When false (the
default), queryProductsWithCache drives the keyword-only fallback:
one backend search per positive OR-group, unioned and deduped, with the full
boolean predicate enforced client-side by fuzzyFilterAst.
Private Readonlymax
Supplier implementation for ScienceLab, a US chemical supplier running on BigCommerce (sciencelab.com). BigCommerce has no matching platform base, so this extends SupplierBase directly.
The on-site search returns only 12 products per page, so instead the whole catalog is pulled from the XML sitemap (one request), product names are recovered from the URL slugs, and matches are ranked locally. The top-
limitproduct pages are then scraped for the real title, price, CAS, formula, and per-size variant pricing (each size priced via the BigCommerce product-attributes AJAX endpoint).Type Param: S
The supplier-specific search item (ScienceLabItem)
Type Param: T
The common Product type that all suppliers map to
Example
Source