Normalized procurement records
Forecast, opportunity, award, supplier, and change records are mapped into documented schemas with stable identifiers and explicit source context.
WebTruffle turns public U.S., EU, and UK tender notices plus U.S., EU, and UK contract awards into normalized opportunity, award, supplier, contract, history, and change records. Start with the free editions; add the sources, history, fields, qualification rules, and delivery your workflow needs.
Need a different recurring dataset? See managed scraping beyond procurement.
Before WebTruffle, across web scraping and quality assurance.
Three free, ungated procurement datasets show the collection, normalization, history, and provenance model in public. Use them directly, or use the documented editions to define the gap a tailored feed must close.
DHS · Energy · Education · Justice
Daily snapshots of planned requirements from four official agency forecast systems, with observed changes and explicit coverage limits.
US federal · EU · UK procurement
Normalized opportunity notices from SAM.gov, TED, Find a Tender, and Contracts Finder, with current, history, change, and market-summary products.
US prime awards · public suppliers
USAspending prime award summaries shaped into award, supplier, and award-recipient products with explicit cumulative-value semantics.
No source-specific collectors, schedules, validation rules, history reconstruction, or ordinary source-change fixes for your team to own. You receive procurement data mapped to an agreed schema while one accountable team operates the feed.
Forecast, opportunity, award, supplier, and change records are mapped into documented schemas with stable identifiers and explicit source context.
Receive the agreed procurement dataset daily, weekly, monthly, or on the collection cadence your workflow needs.
History and change products make amendments, status changes, and newly published records easier to detect and route.
Every feed has clear engineering and QA ownership from scoping through production.
For adjacent public or client-authorized sources, WebTruffle can scope, extract, normalize, validate, deliver, and maintain a custom feed without a scraping platform or contractor chain for your team to manage.
Discuss another data feed→Custom extraction workflows built around your target sources, field schema, navigation paths, and collection volume.
Scheduled web data collection from one or multiple sources for analytics, research, operations, or data products.
Standardized dates, currencies, units, categories, and identifiers mapped into one dependable schema.
Acceptance criteria defined before production, with automated validation and manual QA for accuracy, freshness, and edge cases.
Run monitoring, retries, anomaly alerts, schema-drift detection, and maintenance for routine source changes within the agreed scope.
Receive CSV or JSON files through Microsoft Azure Blob Storage, Amazon S3, or a tailored dashboard.
Procurement intelligence is the primary product. These adjacent use cases apply the same schema, quality, delivery, and maintenance standards to other recurring data needs.
Track catalogs, specifications, sellers, prices, promotions, ratings, availability, and assortment changes.
E-commerce · Retail · Marketplaces
Monitor launches, offers, locations, reviews, listings, and the public signals shaping your market.
Research · Strategy · Monitoring
Aggregate frequently changing listings with standardized fields, timestamps, deduplication, and history.
Real estate · Aggregators · Portals
Collect public job postings for aggregation, workforce analysis, labor-market research, and talent intelligence.
Job boards · HR tech · Research
Build consistent company, location, and directory datasets from fragmented public sources.
Directories · Company data · Enrichment
Deliver public or authorized web data into dashboards, reports, applications, and internal systems on a recurring schedule.
Custom feeds · Cloud delivery · Dashboards
We agree on the sources, schema, acceptance criteria, delivery, and ownership before production begins.
We assess the sources, fields, volume, cadence, delivery, access conditions, intended use, and major risks—and recommend an official API when it is the better route.
For higher-risk work, we produce a representative sample against an agreed schema and acceptance criteria before full production.
We schedule delivery, monitor feed health, investigate failures, and maintain the workflow as sources change within the agreed scope.
Automation catches routine failures; experienced QA challenges the edge cases. Failed records can be flagged, quarantined, retried, or reported according to the project specification.
Required-field coverage
Missing and null-rate checks
Data types and formats
Allowed values and ranges
Duplicate and record-count anomalies
Freshness and delivery timing
Schema-change detection
Source-to-output sampling
Our founders bring complementary experience across the full lifecycle of a managed data feed: building the web software, extracting and structuring the data, and challenging the result before it reaches you.
Sixteen years building production web systems, including six years focused on custom scraping workflows and dependable data extraction pipelines.
Ten years combining hands-on exploratory testing with automated checks to uncover broken assumptions, regressions, and edge cases.
Start with the free public editions and a requirement review. Use a representative paid pilot when the gap needs proof, then move to a managed feed only after the coverage and output are clear.
A focused assessment of the markets, sources, records, fields, cadence, delivery, and gaps behind your procurement use case.
Coverage fit · Key risks · Recommended next step
A one-time setup fee that validates the hardest assumptions with representative procurement records, an agreed schema, and sample output.
No second setup fee when the same agreed scope moves into a managed feed
A recurring procurement feed operated end to end and billed at the end of each service month: collection, validation, monitoring, delivery, and maintenance.
Scheduled delivery · Quality monitoring · Ongoing maintenance
Prices exclude VAT. Final pricing depends on source complexity, frequency, volume, transformations, quality requirements, delivery, and support expectations.
Understand what is free, when a managed procurement feed helps, and where broader managed scraping still fits.
Government procurement intelligence turns public opportunity, award, supplier, and change records into structured evidence for market research, qualification, product features, and operational workflows. The useful layer is not just collecting notices; it is preserving their geography, lifecycle, identifiers, provenance, and material changes.
Yes. The public forecast, tender, and contract-award editions are ungated and include documented downloads and previews. They are designed to be useful on their own and to show the data model before you consider a tailored feed.
A managed feed is useful when the free edition does not cover the sources, history, fields, qualification rules, cadence, or delivery your workflow needs. We scope that gap, validate it with a representative sample when appropriate, and operate only the agreed extension.
Yes. Procurement intelligence is WebTruffle’s primary product, while managed scraping remains available for adjacent public or client-authorized data requirements. We can scope recurring feeds for products, prices, markets, listings, jobs, business directories, and operational sources.
No provider can responsibly promise that. We assess technical feasibility, access conditions, the type of data, intended use, and relevant legal or contractual constraints. We focus on public or client-authorized data and may decline projects that create unacceptable risk.
We agree on a field schema and acceptance criteria, then combine automated checks with manual QA. Checks can cover required fields, data types, allowed values, duplicates, record counts, freshness, schema changes, and samples compared with the source.
Managed plans include monitoring and maintenance for ordinary source changes within the agreed scope. If a source removes the data or changes access rules fundamentally, we assess the impact and agree on the next step with you.
Yes. We can assess an existing scraper or feed, identify reliability and quality gaps, and recommend whether it should be stabilized, rebuilt, or replaced.
After a short scoping conversation, you receive a technical-fit assessment covering the proposed sources, fields, cadence, delivery method, and key risks, plus our recommended next step. We may suggest an official API, a paid pilot, a production scope—or tell you the project is not a good fit.
The paid pilot starts at €750 + VAT and acts as the one-time setup fee for the agreed scope. If that scope moves into a managed feed, there is no second setup charge. The managed service starts at €1,000/month + VAT and is billed at the end of each service month.
Your delivered project data remains yours, subject to applicable source and third-party rights. WebTruffle retains its scraper code, reusable libraries, internal tooling, and technical know-how. At offboarding, we provide the agreed final data export and remove project access according to the contract and retention requirements.
WebTruffle works with customers worldwide. Delivery can be arranged as CSV or JSON files through Microsoft Azure Blob Storage, Amazon S3, or a tailored dashboard, depending on the agreed project scope.
Procurement requests capture the market, geography, intended use, timing, and delivery requirement so the sample is genuinely useful. General managed-data enquiries can still use the shorter contact path.
Need billing, customer support, privacy, security, or something else? Use the general contact form.
Preparing the requirement form
Tell us what the data needs to do.
Loading the relevant public-dataset context and delivery questions…
No generic sales deck. If the project is not a good fit, we will say so.