The graph with a trap — the panel's strongest call
The standing graphing task: a first-hand data table (often a gas, growth or rate context) to be plotted with correct axes (independent variable on x), uniform labelled scales with units,...
The question styles to have ready — each linked to its evidence and to a question in the practice paper.
Built by a six-model AI panel and backtested against the hidden 2025 papers — how we did it.
Built from a six-model AI analysis of every HSC Investigating Science paper, marking guideline and marking-centre feedback report since 2019. First, the housekeeping: the current syllabus continues into 2026 with no replacement announced — 2026 is NOT a final syllabus year, so ignore any "last chance, everything must appear" cramming folklore. These are styles to prepare for, not guarantees: when we backtested this method on the 2025 papers, the examiners kept the skill and twisted the format — so practise the skill chain, not a memorised question.
The standing graphing task: a first-hand data table (often a gas, growth or rate context) to be plotted with correct axes (independent variable on x), uniform labelled scales with units,...
The Module 8 anchor: a graph or table of research spending (government vs corporate, or competing national priorities) — often with a dual-scale or proportion trap — where students...
A described human or animal investigation needing ethical clearance: justify ethics-committee requirements (informed consent, prior animal or in-vitro work, welfare, right of withdrawal)...
a first-hand data table to plot: axes labelled with units, independent variable on the x-axis, uniform scales, then a line or curve of best fit — with exactly one planted trap. Recent traps: an outlier your line must ignore (2024), and a non-linear trend demanding a curve (2025). A gradient or "does this support the hypothesis?" part follows.
a construct-a-graph task has appeared in six of the last seven papers, and all six models call it again — our highest-consensus cluster.
forcing the line through the outlier or the origin; swapping the axes; non-uniform scales.
one graph a week from raw data; before drawing, ask "outlier or curve?" — one of them is almost certainly hiding in there.
5–7 marks: a testable everyday hypothesis (2025 used sleep and heart rate) and a blank page. Numbered steps, independent and dependent variables, a separate control group — not the same thing as your controlled variables — a quantitative measurement, repetition, and a safety step matched to the actual hazard.
a method item has appeared in six of seven papers, and four of six models predict the full write-a-method form again.
confusing the experimental control with controlled variables; offering repetition as a validity fix; describing measurement that yields no numbers.
a described study with deliberate faults — no control group, a confounded variable, a tiny skewed sample — then: identify the faults and justify two or three modifications, each tied to the one criterion it improves (validity, reliability or accuracy).
the 2024 Q27(c) / 2025 Q31(b) lineage, called by five of six models. It is the committee's most reliable discriminator because of the next line.
answering about the wrong criterion. "Evaluate the validity" answered with repetition talk scores almost nothing. Name the criterion in every sentence.
a product claims a number ("mass up 20% in three weeks", 2025); trial data lets you calculate the actual change. Compute it, compare it with the claimed figure, THEN judge — and a chained part asks what the design was missing (usually the untreated control).
the 2022 cholesterol / 2023 moisturiser / 2025 fertiliser template, called by five models; feedback in 2023 and 2025 explicitly demands the numeric comparison before the verdict.
percentage-change calculations inside claim scenarios, ending every answer with a supported yes/no on the claim.
repeated readings from an analogue and a digital instrument against a known true value, with one outlier planted. Exclude the outlier, average, judge accuracy by closeness to the TRUE value (never by comparing the devices to each other), then name the poorer device's error: constant offset = systematic, scatter = random — each with a procedural cause.
averaging the outlier in; "human error" as an answer (it never scores); precision confused with accuracy. The panel expects the measurand to rotate off 2025's temperature — mass, volume or pressure are due.
The panel's one confident rested call: the full science–technology "continuous cycle" extended response closed the paper with 7 marks in 2023 and 8 marks in 2025, and the panel strongly expects it NOT to close 2026. Know the content — a short 3–5 mark single-link question (X-ray diffraction → DNA, radioactivity → nuclear medicine, LHC → Higgs) is still likely — but don't build your revision around another full-cycle essay.
We backtested this exact method by having the same six models predict the 2025 Chemistry and Maths Extension 1 papers blind, then scoring them against the real exams. The high-confidence topic calls went 37/37 — but at question level only about half of the specific predictions recognisably appeared, and the examiners inverted formats, migrated questions to multiple choice, and broke streaks the whole panel trusted. Investigating Science itself was not in that backtest, and its papers lean even harder on rotating surface contexts around stable skill chains. So treat every scenario in this guide as a costume: prepare the skill chains — graph conventions, the validity/reliability/accuracy distinction, numbers-before-verdicts — and you're covered whatever the 2026 paper dresses them in.
The standing graphing task: a first-hand data table (often a gas, growth or rate context) to be plotted with correct axes (independent variable on x), uniform labelled scales with units, and a line or curve of best fit — with exactly one planted trap: an outlier the line must ignore, or a non-linear trend demanding a curved fit (2025 Q34's escalation). Chained to a gradient, relationship or hypothesis-support interpretation part.
Marker-feedback lineage: Marking feedback 2019, 2020, 2022, 2023, 2024 Q32 (outlier), 2025 Q34(b) (curved fit)
In the practice paper: Q34
The 5–7 mark 'write a valid method' item: a testable hypothesis in an everyday, physiological or kitchen-bench setting (2025 sleep/heart-rate, 2024 gas syringe lineage); students sequence numbered steps naming the independent and dependent variables, a separate experimental control (distinct from controlled variables), quantitative measurement of the dependent variable, repetition, and a hazard-matched safety step. Top band requires validity features and reliability features explicitly.
Marker-feedback lineage: Validity/reliability confusion flagged every year 2019–2025; 2024 Q28 and 2025 Q26 demanded explicit validity design
In the practice paper: Q26
A described investigation with planted faults — missing control group, confounded variables, small or skewed sample, single trial — and a chained demand: identify the faults, then justify two or three modifications, each explicitly linked to the ONE integrity criterion it improves (validity, reliability or accuracy), without changing the inquiry question. The 2024 Q27(c) / 2025 Q31(b) lineage; answers built on the wrong criterion score near zero.
Marker-feedback lineage: Marking feedback 2020 Q25, 2023 Q32(a), 2024 Q27(c), 2025 Q31(b)
In the practice paper: Q32
A consumer or wellness product states a quantified claim (the 2025 Fertiliser Z 'mass up 20% in three weeks' template); supplied trial results let students calculate the actual change, compare it to the claimed figure before judging, then critique the missing control, confound or unrepresentative sample in a chained part. The numeric comparison must come before the verdict.
Marker-feedback lineage: Marking feedback 2022 (cholesterol), 2023 Q32 (moisturiser), 2025 Q31 (fertiliser): students must use the numbers
In the practice paper: Q31
A repeated-trials table for two instruments against a stated true value, seeded with one outlier: exclude it before averaging, judge which device is more accurate citing numerical closeness to the TRUE value (not to each other), and attribute the poorer device's constant offset to systematic error and its scatter to random error, with a procedural cause. Grok specifically predicts the measurand changes from 2025's temperature (mass, volume or pressure due).
Marker-feedback lineage: Marking feedback 2020 Q13–15, 2022 Q22, 2023 Q24, 2025 Q27 (outlier + true value), 2025 Q35 (error types)
In the practice paper: Q27
A wellness or therapy claim (copper bracelets 2024, green tea 2021 lineage) to be tested properly: describe a trial specifying what the placebo is, who is blinded and how allocation is concealed, random assignment, the control group and the measured outcome — then justify how EACH element removes participant expectation or researcher bias. Definitions without deployment in the scenario cap in the low bands.
Marker-feedback lineage: Marking feedback 2021 Q29, 2024 Q35: students defined placebo/double-blind but could not deploy them
In the practice paper: Q33
The Module 8 anchor: a graph or table of research spending (government vs corporate, or competing national priorities) — often with a dual-scale or proportion trap — where students describe the trend QUOTING values, then analyse how the funding source steers project choice, duration and reported outcomes, closing with a judgement. Named bodies (ARC, NHMRC, a named firm) required; 'the government' scores thin.
Marker-feedback lineage: Marking feedback 2019 Q34, 2022 Q33, 2024 Q34, 2025 Q20/Q30: quote the data, name the body
In the practice paper: Q35
A described human or animal investigation needing ethical clearance: justify ethics-committee requirements (informed consent, prior animal or in-vitro work, welfare, right of withdrawal) against the scenario's specific risk — kept distinct from validity critique — and/or name a specific code (Nuremberg Code, Declaration of Helsinki/Istanbul, NSW Animal Research Act) stating what it PERMITS or FORBIDS, with a judgement of effectiveness. After 2025's describe-TWO-codes (Q33), the panel expects the apply-and-evaluate variant.
Marker-feedback lineage: Marking feedback 2020 Q25(b), 2023 Q26, 2024 Q27(a), 2025 Q33
In the practice paper: Q28
A named syllabus scientist — Marshall and Warren, Priestley, van Helmont, Spencer, Eratosthenes, Doppler, Jenner — with the demand being HOW the investigation began with observation and departed from the linear model (accidental redirection, self-experimentation, a sample of one). Rotation logic splits the panel: Marshall/Warren carried 2025, van Helmont and Spencer 2024, so fable calls Doppler or Priestley due; grok backs an Eratosthenes justify-the-method.
Marker-feedback lineage: General feedback 2022–2025 verbatim: 'recognise the importance of the work of scientists named within the syllabus'; 2021 Q24, 2025 Q23
In the practice paper: Q21
Aboriginal and Torres Strait Islander Peoples' knowledge of a NAMED plant (spinifex, Kakadu plum, tea-tree, gumby gumby) underpinning a commercial or medical product: explain why that knowledge is valued by the scientific community AND the benefit-sharing or intellectual-property obligation an ethical partnership owes in return. Ran every year 2022–2025; after 2025's spinifex-latex ethics angle (Q29), grok predicts the swing back to use-and-value with a different plant.
Marker-feedback lineage: Marking feedback 2019 Q27, 2023 Q23, 2024 Q24, 2025 Q29
In the practice paper: Q23
A stimulus asserting a relationship — two series rising together, or a claimed health/cognitive benefit — where students define the correlation, explain why causation is not established, name a plausible confounding variable, and outline the controlled evidence that could test causation. Opus expects a vehicle other than the Mozart effect after 2025 Q25.
Marker-feedback lineage: Marking feedback 2019 Q25(d), 2023 Q21, 2025 Q25
In the practice paper: Q25
A sponsor's or outlet's presentation that flatters its position: a truncated vertical axis, projections replacing actuals, mismatched dual scales (2024 Q23 / 2020 Q24 mould) — or a headline/caption misusing 'theory', 'hypothesis', 'law' or 'proven' set against a journal extract. Students name the specific distorting feature or term, quote it, and explain how the presentation shifts public perception, ending with a verdict.
Marker-feedback lineage: Marking feedback 2020 Q24, 2021 Q28(b), 2023 Q25(b), 2024 Q23
In the practice paper: Q29
Competing-viewpoints closer: 2–3 quoted students/scientists on a Mod 7/8 issue, verdict must reference every statement — three straight years 2023–2025 (fable 0.50, gpt-5.6-sol 0.48, opus 0.42)
Continuous-cycle closer RESTED is the panel's strongest structural call: after 7-mark 2023 Q36 and 8-mark 2025 Q36, grok puts P(no cycle closer) at 0.70 and gpt-5.6-sol 0.56; deepseek is the contrarian, keeping a 7-mark radioactivity cycle (45% chance)
Single-link science–technology item survives the rest: one 3–5 mark named chain (radioactivity→atomic theory→nuclear medicine, X-ray diffraction→DNA, LHC→Higgs) — all six models, mean 0.52 (gemini 0.75, opus radioactivity bold 0.35: dormant since 2020 Q26)
Pseudoscience vehicle rotation: after astrology (2025) and numerology (2024), iridology is due — fable bold 0.35 with practitioner-inconsistency data, gpt-5.6-sol 0.55, gemini-3.1-pro 0.55, deepseek 0.45
Conflict-of-interest item requiring an example from a DIFFERENT industry (fable 0.70, opus 0.55) — but grok explicitly rests it after 2025 Q28's Brand X
Halo-effect MC with Hawthorne as chief distractor is near-annual (grok 0.75); deepseek's bold inversion: Hawthorne as the CORRECT answer in a workplace scenario (35% chance)
Sequencing MC (order the steps of a named investigation) ran 2023–2025; fable 0.85, opus 0.65
Depth-study own-investigation recall ('state the claim you tested, outline your procedure') returns after 2023 Q35 (fable bold 0.30, deepseek structural 0.60)
Doppler's first 4+ mark treatment since 2020: frequency-vs-time record of a moving source (fable bold 0.35)
Student-report conventions evaluation (structure AND language), unused since 2023 Q30 (grok bold 0.42, opus 0.35)
Gas-behaviour context for the eighth consecutive year, MC or practical (fable 0.80)
gemini-3.1-pro's contrarian rested call: Mod 6 applications/instruments rested entirely (P(examined) 0.35) — outvoted 5–1 by the panel
How likely each topic is to appear this year.
Chance of a big question (4+ marks) here: 94%
Question types predicted here extended response ×8 stimulus based ×2 short answer ×2 practical analysis ×2 multiple choice ×2
What each model expects
Chance of a big question (4+ marks) here: 84%
Question types predicted here short answer ×6 extended response ×4 stimulus based ×3 multiple choice ×3
What each model expects
Chance of a big question (4+ marks) here: 81%
Question types predicted here extended response ×8 short answer ×4 stimulus based ×3 multiple choice ×1
What each model expects
Chance of a big question (4+ marks) here: 63%
Question types predicted here short answer ×9 extended response ×3 multiple choice ×2
What each model expects
Chance of a big question (4+ marks) here: 64%
Question types predicted here short answer ×6 extended response ×3 multiple choice ×3 stimulus based ×2
What each model expects
Chance of a big question (4+ marks) here: 60%
Question types predicted here short answer ×7 extended response ×3 multiple choice ×3
What each model expects
Chance of a big question (4+ marks) here: 65%
Question types predicted here short answer ×8 stimulus based ×4 multiple choice ×2
What each model expects
Chance of a big question (4+ marks) here: 70%
Question types predicted here short answer ×5 multiple choice ×5 practical analysis ×3 stimulus based ×1
What each model expects
How likely each topic is to appear. Open a topic for the question types to practise there.
100 marks · 36 questions
Every question is traceable to the consensus prediction behind it — open the web version and each question carries a “why this question” link into the evidence. All questions are original Intuition compositions in NESA style.
Intu AI
Intu AI builds unlimited practice questions for Investigating Science in these styles, marks your working, and explains what you missed — aligned to your syllabus.
Published Aug 2026, before the exams. In November 2026 we score these predictions publicly against the real paper — per-model calibration and question-level hit rates, the same harness as the 2025 backtest. How we did it.