Skip to content
Clothsy AI Talk to us

Chapter 4The programme7 min readVersion 1.2 · 30 September 2026

R&D roadmap

Eighteen months in seven phases, fifteen milestones and four papers, with the risks named up front.

The programme runs for eighteen months in seven phases. The image track fills most of the first year; the video track starts in month 4 with its capture protocol and becomes the main effort from month 9. Month 1 is indicatively October 2026. If funding starts later, the months shift and the venue plan moves to the next cycle of each venue.

4.1Phases

Table 28 Programme phases, deliverables and exit criteria
PhaseMonthsKey activitiesDeliverablesExit criterion
P0 Foundations1Company registration and DPIIT recognition; evaluation harness; consent, annotation and data protocols; licence register; cloud credit applicationsHarness running on at least six systems; approved protocolsHarness reproduces published numbers within tolerance on a public split
P1 Pilot2Pilot shoot of 20 garments across the garment families; annotation guide; pilot auditPilot set; inter-rater agreement reportKrippendorff's alpha of at least 0.6 on garment attributes
P2 Benchmark3 to 6Main collection and annotation; audit of at least ten systems; human study; Paper 1Benchmark v1 frozen; Paper 1 on arXiv and submittedBenchmark meets its size, garment-family and strata targets
P3 Image model3 to 9Data engine; synthetic triplets; LoRA stages; multi-view conditioning; ablations; preference tuningImage model; ablation reportFirst LoRA beats the current production model on pilot garment fidelity
P4 Distil and transfer9 to 12Distillation; serving tests; serving-pipeline clean-up; Paper 2; image-track releasesDistilled image model; Paper 2 submitted; toolkit and benchmark releasedDistilled model within twice 6.5 seconds on an L40S with no significant loss in preference
P5 Video model4 to 14Video capture protocol and pilot shoot; main shoots; video data engine; offline video model; Paper 3Video training set v1; offline video model; Paper 3 submittedOffline model beats open video baselines on garment fidelity over time on the hold-out
P6 Live12 to 18Causal conversion and distillation; streaming servers; live pilot behind the existing live page; Paper 4Live model; pilot report; Paper 4 submittedAt least 15 fps at 512p on one GPU, first frame under 2 s, cost under a tenth of the rented engine

4.2Schedule and milestones

Milestones
1234·56789·101112131415
WP0 Foundations: entity, harness, protocols
WP1 Pilot shoot and annotation guide
WP1 Benchmark collection and annotation
WP1 Baseline audit and human study
WP5 Paper 1 writing and submission
WP2 Data engine and synthetic triplets
WP3 LoRA training stages and ablations
WP3 Multi-view conditioning experiments
WP4 Preference tuning and distillation
WP5 Paper 2 writing and submission
WP4 Serving integration in the product
WP6 Video capture protocol and pilot shoot
WP6 Video shoots and synthetic video pairs
WP6 Offline video model training
WP5 Paper 3 writing and submission
WP7 Causal conversion and distillation
WP7 Streaming servers and live pilot
WP5 Paper 4, releases and final report
Programme month
123456789101112131415161718
  • Foundations
  • Image track
  • Video track
  • Live track
  • Papers
Figure 3Programme schedule by work package, with milestones MS1 to MS15. Month 1 is indicatively October 2026.
Table 29 Milestones
IDMilestoneMonthEvidence
MS1Evaluation harness running on at least six systems1Harness report
MS2Pilot set and inter-rater agreement report2Pilot report
MS3Benchmark v1 frozen5Datasheet and data card
MS4Paper 1 posted to arXiv and submitted6Submission record
MS5Video pilot: capture protocol tested and first 200 consented clips6Pilot report
MS6First LoRA beats the current production model on pilot garment fidelity7Evaluation report
MS7Image ablations complete8Ablation report
MS8Paper 2 submitted10Submission record
MS9Distilled image model meets the latency target11Speed laboratory report
MS10Video training set v1: 7,000 real clips and 20,000 verified synthetic pairs11Data card and provenance ledger
MS11Image-track releases and interim report12Repositories and interim report
MS12Offline video model beats open video baselines on the hold-out13Evaluation report
MS13Paper 3 submitted14Submission record
MS14Live prototype: at least 15 fps at 512p on one GPU, first frame under 2 s16Live evaluation report
MS15Live pilot with shoppers, Paper 4 submitted and final report18Pilot report, submission record and final report

4.3Publication calendar

Table 30 Target venues and indicative deadlines
VenueIndicative deadlineStatus on 30 September 2026Target
CVPR 2027, main conference [63]16 November 2026ConfirmedToo early for any paper
CVPR 2027 workshopsAround March 2027Not yet announcedPaper 1
ACM Multimedia 2027Around late March to April 2027Not yet announcedPaper 1 or Paper 2
NeurIPS 2027 Evaluations and Datasets [64]Around May 2027Not yet announcedPaper 1, extended
BMVC 2027Around late May 2027Not yet announcedPaper 2 if results are early
WACV 2028Around June and August 2027Not yet announcedPaper 2 or the cost report
CVPR 2028Around mid-November 2027Not yet announcedPaper 3
SIGGRAPH 2028Around January 2028Not yet announcedPaper 4 if results are early
ECCV 2028Around March 2028Not yet announcedPaper 4

Deadlines marked 'around' follow each venue's usual cycle and will be confirmed when calls open. These venues allow arXiv preprints if the submitted paper is anonymous.

4.4Governance and ways of working

  • A weekly research meeting with a shared agenda, and a monthly steering review with the founders.
  • One code repository, one experiment tracker and versioned configurations, so every reported number can be reproduced.
  • A named data steward who owns the provenance ledger, consent records and release decisions.
  • An authorship and contribution policy agreed in writing before work starts.
  • Quarterly progress reports to funders against the milestones and the key performance indicators.

4.5Risk register

Table 31 Risks and mitigations
RiskLikelihoodImpactMitigation
A licence contaminates the modelMediumHighApache-2.0 or MIT bases, teachers and annotation tools only; licence register and provenance ledger; legal review before any release
Consent or privacy failureLowHighWritten consent and releases covering AI training and synthetic derivatives; no scraping; anonymised public release
Another group publishes firstMediumMediumPost to arXiv early; keep the focus on garment coverage, fairness strata and garment-faithful live try-on
Data collection is slower than plannedMediumMediumPilot first; catalogue partners in parallel; a 1,000-pair minimum image benchmark as fallback
Diverse subjects and garments are hard to source from one countryMediumMediumCatalogue, stock and footage partners in several regions; partner shoots abroad; coverage reported honestly for every stratum
Video data is costly and slowHighMediumPilot shoot before Tier A, Tier A before Tier B; shooting in India lowers cost; licensed footage covers the unpaired stage
Live garment accuracy falls shortMediumHighShip offline video first; keep the rented engine for pilots; report the gap against the hold-out honestly
Serving cost of large modelsHighMediumDistil the image model into a 4B student and the video model into a 1.3B live student; keep the current production model until the student wins
GPU capacity for training and live servingMediumMediumCloud credits across providers; the portable worker container already built; one GPU per live stream by design
Live video is misused to fake someone's appearanceMediumHighThe guardrail plan: age and consent checks at session start, sampled output moderation, session limits and labels
Volunteer attritionMediumMediumFunded fellowships, scoped tasks and clear authorship

Sources in this chapter

  1. [63]CVPR 2027. Key dates. cvpr.thecvf.com/Conferences/2027/Dates
  2. [64]NeurIPS 2026. Call for Evaluations and Datasets. neurips.cc/Conferences/2026/CallForEvaluationsDatasets