๐ฅ Roast My Pick ยท SIH26063
Integrated Polar Science Outreach, Knowledge Repository and Media Dissemination Portal
Ministry of Earth Sciences (MoES)
Bold. Let us find out precisely how bold, in the order a panel will find out.
Proceed with caution. The corpus is real and the build is safe, but an archive with an AI post generator has a very low ceiling โ take it only if you will make grounded, citation-locked generation and multimodal search the substance rather than shipping a CMS. Roughly 95โ220 teams are expected to go here.
The receipts
Every red flag on this statement, in full. These are the four places it bites.
Exhibit A
An archive plus a text generator is among the least technically distinctive products you can build, and a panel will have seen several retrieval-and-generate submissions before yours
It gets worse
Generated science communication that subtly misstates a finding is worse than no post at all, so the editorial review queue is a requirement rather than a nicety and should be visible in your demo
Still reading?
Six media types is deceptively broad โ video ingestion, transcription and indexing alone can consume the whole event if you let it
And the finisher
There is no measure of good outreach in the statement, so any claim that your generated content is effective rests on your own judgement and a judge may simply disagree
The damage report
Every score this statement earned, and what each one actually costs you.
Feasibility
4/5Actually buildable, which on this slate is rarer than it sounds. Do not squander it on scope.
Indian Antarctic expedition reports, station documentation and imagery are published openly enough to assemble a genuine corpus, and repository search plus grounded text generation are mature and well-supported โ nothing here depends on data you cannot obtain.
Innovation scope
4/5There is something genuinely new here. Do not bury it under another dashboard.
At 213 characters the statement names what the portal holds and that it should generate content, and prescribes nothing about how โ the retrieval design, the generation grounding and the editorial workflow are all yours.
Clarity
2/5Nobody is sure what is being asked, quite possibly including the people who asked it.
Short, but unlike the neighbouring polar statements it does at least enumerate the content types and identify two distinct functions in archiving and generating, so you are not starting from nothing โ though there is no audience, no scale and no measure of what good output would be.
Acceptance potential
2/5The numbers do not like you. Bring something the numbers cannot see.
A content repository with an AI caption generator is close to the lowest-ceiling software shape available โ the archive half is a CMS and the generation half is an LLM call, so there is very little for a judge to reward beyond execution and the outreach framing does not create technical substance.
Effort
HeavyHeavy. Somebody on this team is not sleeping in week three. Pick who, on purpose.
Ingestion across six media types, metadata extraction, a search layer, a generation pipeline, an editorial review queue and a scheduling interface is six components, and handling video and imagery properly is more work than the text side.
Demo-ability
MediumDemoable, if you rehearse it. Nobody rehearses it.
Search over a real archive works reliably and grounded generation with highlighted sources is a decent moment, but a content portal generating social posts is a familiar sight in 2026 and nothing here will surprise a panel.
Data
None suppliedNo dataset comes with this one, so every accuracy figure you quote is a number about labels you invented.
Nothing is provided with the statement. You are sourcing, cleaning and labelling it yourself, and that work is invisible in the demo but very visible in the questions.
The demo they will have already seen
Somewhere around 95โ220 teams are heading here, and the description is doing the choosing for most of them. They will read the same brief, reach the same architecture, and build a version of the same demo you are planning. Being correct is the floor. If your five minutes could be swapped with the team before you and nobody in the room would notice, you have not picked badly โ you have built predictably, which costs exactly the same and hurts more.
What survives
The ground worth standing on when the questions start.
- Indian polar expedition material is genuinely published and downloadable, so unlike the neighbouring NCPOR statements your corpus is real rather than invented
- Grounding generated posts to specific passages in a source report is a real and demonstrable discipline, and a portal that refuses to write anything it cannot cite is a defensible differentiator over freeform generation
- Semantic search across photographs and video, not just text, is the harder and more interesting half and almost no competing team will attempt it
None of that means do not pick it. It means do not walk into that room having heard any of this for the first time from a judge.
The framing is a joke. The findings are not โ they are the same analysis on the statement page, and every line above is attached to a score or a fact in the record. It is one opinion with its reasoning attached, so argue with it before you trust it.