Logical Reasoning

LSAT Language Strength: How Strong Can an Answer Be?

The two modes, the words that change everything, and how to know when an answer goes too far.

By Jeremy LaRusso · Updated September 2026

There is a whole category of LSAT question you can understand perfectly and still get wrong. You read the stimulus correctly. You find the conclusion. You understand the evidence supporting it. Then you choose an answer that says slightly more than the argument allows, and you lose the point because of a single word.

Language strength is one of the most reliable eliminators on the Logical Reasoning section, and it takes students a surprisingly long time to trust, because it sounds at first like grammar rather than logic.

It is logic. Some is different from most. May is different from will. Helped cause is different from caused. More effective makes a comparison that effective does not.

This is not a difficult idea, only an unfamiliar one. The phrase gets used long before anyone defines it — students are told to watch the strength of the language without being told what that means. So this article starts from nothing.

First: what does “language strength” mean?

Language strength is how much a sentence commits itself to. Compare these:

Some students passed.
Most students passed.
Every student passed.

Each one commits to more than the one above it. The same ladder runs through every other kind of word:

The treatment may help.
The treatment probably helps.
The treatment always helps.

A stronger statement tells you more. It also needs more support. And the relationship only runs one way: if every member of a group has a trait, then most do and some do. If some do, you know nothing about whether most do. If most do, you know nothing about whether all do.

One thing trips almost everyone at first. In conversation, some implies “not all” and most implies “not quite all.” On the LSAT neither carries that upper limit. “Most students passed” does not tell you that anybody failed — if all 100 passed, it is still true that most passed. “Some students passed” does not tell you that anybody failed either. Most permits all, and some permits all. Students lose points assuming the conversational limit is there.

The language bank

These are the working LSAT meanings and minimum commitments — not impressions about how the words sound in ordinary conversation. What matters is the minimum each expression commits the author to. Skim this, then come back to it as the examples below use it.

The categories matter as much as the entries, because these words fail in different ways. A quantity that was never established, a comparison where the stimulus only ever discussed one thing, and a probability where the stimulus only said something was possible are three different wrong answers — and all three are extremely common.

Quantity — how much of the group
all / every / eachevery member of the group, no exceptions
none / nozero members of the group
most / a majority / more than halfmore than 50 percent. It does not mean “most but not all” — if all of them did, then most of them did.
manya substantial but unspecified number. Formally it guarantees only at least some, and it never guarantees a majority — do not convert it into most. It is worth flagging anyway, because it reads like a majority, and that gap between what it means and what it sounds like is exactly what the test uses it for.
some / at least oneone or more. It does not mean “some but not all.” It has no upper bound.
fewa small number. Still a some.
Frequency — how often, which is not the same as how many
always / neverevery time, or no time. Universal.
usually / generally / typically / ordinarilymore often than not — the frequency analogue of most. Do not assume anything stronger than that.
often / frequentlyrepeatedly, with no proportion attached. Do not convert it into “most of the time.”
sometimes / occasionallyon at least some occasions
Likelihood — a probability has to be established before an answer can use one
certain / must / will / cannot fail toguaranteed
probably / likely / more likely than notbetter than even odds
more likely than Xcareful: this compares two probabilities. Something can be more likely than the alternative without being more likely than not.
tends tocontext-sensitive. Often it works like generally or usually; sometimes it describes a direction instead, as in “as X rises, Y tends to fall.” Read the sentence rather than assigning the phrase a fixed percentage.
may / might / could / canpossible. Says nothing at all about how likely. Something can be possible and wildly improbable.
Degree and causation — how much of the work a factor did
solely / entirely / exclusivelynothing else contributes
causes / is responsible forasserts a causal relationship
helps / contributes to / plays a role inasserts some contribution. Not that the factor was sufficient, and not that it was the main one.
associated with / linked to / correlated withasserts a relationship of unspecified kind. Never causation on its own.
consistent withthe evidence does not contradict the claim. Far weaker than showing the claim is true.
in part / partiallysome contribution, not the whole of it
up to 40 percent / as many as 40 percenta maximum, not a minimum. “Up to 40 percent” is satisfied by zero, and does not mean “about 40 percent.”
appears / suggests / seemsreports an impression rather than asserting
Conditionals — strong, but direction matters as much as strength
if A, then BA is sufficient for B
A only if BB is necessary for A
only A are Bif something is B, it must be A
unlessP unless Q = if not Q, then P — and equivalently, if not P, then Q. Write the translation out; the compressed description is where people go wrong.
if and only ifboth directions at once
Comparatives — the comparison itself needs support
more / less / fewer / greatercompares at least two things
better / worse / superior / inferiorestablishes a ranking
higher / lower / closer / farthercompares on some scale
as much as / the same asasserts equality, which still requires two things

Conditionals are the one category where labeling a word “strong” and moving on will actively hurt you. If, only and only if all establish rules, but they point in opposite directions. Diagram them; do not just note that they are forceful.

Comparatives are dangerous for a different reason. An answer can invent an entirely new comparison while every noun in it appeared in the stimulus. If the stimulus told you experienced attorneys are effective negotiators, it did not tell you they are more effective than inexperienced ones. The second sentence is not a more confident version of the first. It introduced a second group.

The two modes

Now that you know what the words mean, the question is how strong the correct answer is allowed to be. That depends entirely on which of the two modes you are in.

Mode 1. The stimulus gives you the information. Your job is to draw an inference, make a judgment, or describe the reasoning, based on what you were given.

Mode 2. You are the one introducing the information. The answer choice enters the argument as a new fact, and you are told to treat it as true.

How you learn this matters. Most students learn it as two lists of question types, memorize the lists, and then stall the first time they meet a stem worded in a way their list did not anticipate. The test is not the list. The test is the rule: if the answer choice is not entering the argument as a new fact, you are in Mode 1.

The lists follow from the rule rather than replacing it. Mode 2 is the short one:

  • Strengthen and Weaken
  • Sufficient Assumption
  • Resolve the Paradox / Explain the Discrepancy
  • Principle questions where you supply the principle that justifies the conclusion
  • Principle questions where you supply new information to justify an application of a principle
  • Most Logically Completes, but only when the word before the blank is a support indicator (since, because, for, as)

Everything else is Mode 1: Must Be True and Inference in all its stem variants, Most Strongly Supported, Most Strongly Suggested, Must Be False and Cannot Be True, Most Logically Completes when the word before the blank is a conclusion indicator, Necessary Assumption, Flaw, Main Conclusion, Role of a Statement, Method of Reasoning, Point at Issue, Parallel Reasoning, Flawed Parallel, the three principle tasks where a principle is already in play, Misinterpretation, Evaluate the Argument, Illustration and Generalization — and in Reading Comprehension, Main Point, Author’s Opinion, another party’s opinion, According to the Passage, Most Strongly Suggests, Definition, Purpose, Value Most Highly, and the hypothetical application questions. Strengthen and Weaken appear in Reading Comprehension too, and they are Mode 2 there exactly as they are in Logical Reasoning.

EXCEPT and LEAST variants do not change any of this. A Strengthen EXCEPT question is still in the Strengthen family; the stem is simply asking which answer fails to do the normal job.

Mode 1: the stimulus is the ceiling

In Mode 1 the answer cannot make a claim stronger than the stimulus supports. I tell students to hold one phrase in their head:

The stimulus sets an upper bound: the answer cannot claim more than the information provided supports. Correct-answer strength ≤ support provided by the stimulus.

Staying under the bound is necessary. It is not automatically sufficient, and what else is required depends on the job:

On inference questions — Must Be True, Most Strongly Supported, Necessary Assumption — a weaker statement can be correct. If the stimulus establishes that every physician completed the training, an answer saying some did is entailed by it and can be the credited response.

On description and matching questions — Main Conclusion, Method of Reasoning, Role of a Statement, Point at Issue, Parallel Reasoning — the answer has to preserve the meaning or the structure as well. The stem asks which answer most accurately expresses the conclusion, and “some physicians completed the training” does not accurately express a conclusion that said every one of them did, even though it follows from it.

The upper bound does the eliminating. If the stimulus says something may occur, you cannot infer that it will. If the stimulus says some members of a group have a trait, you cannot infer that most do. If the stimulus never compared A with B, you cannot infer that A is more effective, more common, or more anything than B. Those answers are gone before you weigh anything else.

Example

Stimulus: Heavy rainfall in spring may increase the mosquito population by midsummer.

Answer choice: Heavy spring rainfall will produce more mosquitoes by midsummer than a dry spring would.

Two separate problems, and either one is fatal. May became will, turning a possibility into a guarantee. And more … than a dry spring is a comparison the stimulus never made — it described one condition, not two.

An important qualification: judge the whole claim, not the loudest word in it

This is where a mechanical reading of the rule breaks, and it is worth getting right, because a careful student will otherwise think the system contradicts itself.

Language strength applies to what the answer choice asserts as a whole — not to every strong-looking word sitting inside it. Consider a Flaw answer:

Example

Answer choice: The argument overlooks the possibility that every dentist included in the survey works for the company that manufactures the toothpaste.

Every is as strong as language gets. But the answer does not assert that every surveyed dentist works for the company. It asserts that the argument failed to rule that out — a far weaker claim, and one that can be fully supported by the structure of the stimulus.

So you cannot do language strength by circling forceful words. You have to see the scope they sit inside. Overlooks the possibility that, fails to consider whether, does not establish that — each of those wraps whatever follows and weakens it.

This is also the reason Flaw is a Mode 1 question type even though its answers sound so absolute. A flaw answer is a claim about the argument, and the argument is something you were handed in full. Either it addressed the sample size or it did not; you settle that by reading the page, not by adding a fact about dentists. The same holds for Method of Reasoning, Role of a Statement, Point at Issue and Parallel Reasoning — all of them describe the argument rather than the world, and all of them can carry forceful language without breaking the ceiling.

The cleanest way to see the difference is to write the same underlying problem in each mode.

The same problem, in each mode

Flaw, Mode 1: The argument fails to establish that enough dentists were surveyed to support a conclusion about dentists generally.

Weakener, Mode 2: Only eight dentists were surveyed.

The first describes a hole in the reasoning and is checked against the stimulus. The second asserts a new fact about the world, and you are told to assume it is true. Same underlying problem, two completely different jobs.

Why Necessary Assumption is Mode 1

This is one of the classifications I most often see taught incorrectly.

A necessary assumption is usually not stated in the stimulus, which makes it tempting to think you are adding it. You are not. A necessary assumption is something the argument was already relying on for its reasoning to work. Your job is to identify that commitment, not to supply it — which is exactly why the Negation Test works. If denying an answer destroys the argument, that tells you the argument needed it all along.

The strength behavior settles it.

Example

Suppose an argument requires only that some members of the committee knew about the problem.

An answer saying every member knew would certainly help the argument. But it is not necessary, because the argument needs only one. The stronger answer demands more than the argument requires, and a Necessary Assumption answer that demands more than the argument requires is wrong.

Sufficient Assumption runs the opposite way: your job there is to add enough to guarantee the conclusion, so strength costs you nothing and is often exactly what you need. Those two question types look like siblings and behave like opposites, and the direction of the strength is what tells them apart. A ceiling set by what you were handed is the defining behavior of Mode 1.

It matters because the wrong version is teachable. A national prep course I took taught that every Sufficient Assumption question is a mismatched concept. It is not — a great many are enabling assumptions, where the gap is not a missing link between two terms but an unstated condition the argument needs in order to work at all. A student taught the false version will hunt for a mismatch, fail to find one, and conclude they are bad at the question type. They are not. They were taught something wrong.

This article also revises a rule of my own that I had been stating more simply than it deserved. Getting it right first time is not the standard worth holding. Not leaving a rule half-reasoned because the simple version is easier to teach is.

Mode 2: use exactly as much strength as the job requires

The standard advice for Mode 2 is “look for strong language.” That is a useful beginner heuristic. It is not the rule.

Use as much strength as the logical job requires. Forceful wording is not worth anything by itself; what counts is the effect on the argument.

On Sufficient Assumption there is a threshold: the answer has to guarantee the conclusion. Clear it and you are done, and extra strength costs nothing.

On Strengthen and Weaken there is no fixed sufficiency threshold. An answer only has to move the argument in the required direction — but because the stem asks which answer does so most, you still compare its impact against the other choices. Two answers can both help; you are choosing the one that moves the argument further.

Often the job does require real force.

Strong language, because the job calls for it

Stimulus: Researchers surveyed 200 UT Law students about whether they preferred vanilla or chocolate ice cream. Sixty-eight percent chose vanilla. So most UT Law students prefer vanilla to chocolate.

Strengthener: Every UT Law student had an equal probability of being selected for the survey.

Every student. Equal probability. That is very strong language, and it is fine — you are in Mode 2, the answer enters as new information, and it directly addresses the worry that 200 respondents might not represent the school. A weak answer would leave that worry open. “Some students who prefer vanilla were surveyed” does nothing at all.

But strong is not automatically better. Sometimes one very narrow fact is devastating.

Weak language, and it finishes the job

Conclusion: No person has ever set foot on Saint Alder Island.

Weakener: At least one sailor from an 1802 expedition went ashore on Saint Alder.

One person. One time. Two centuries ago. It does not establish that hundreds visited, or that anyone visited recently, or that anyone settled there. None of that matters, because the conclusion said no person, ever.

To contradict “no one has ever done this,” you do not have to show that everyone has. You only have to show that one person did.

That is what it takes to contradict a claim about every case: one case going the other way. Universal conclusions are the most fragile things on the test for exactly that reason. An author who writes never has handed you a target a single counterexample destroys.

Move the quantity word and the requirement changes

Example

Conclusion: Most of the plant species on Saint Alder grow nowhere else on Earth.

Learning that one Saint Alder species also grows on the mainland does not contradict this. Neither does learning that two do. A most claim permits exceptions.

Be careful about what that does and does not mean. It does not mean a Weakener has to contain a majority claim of its own. Weakening is making a conclusion less likely, not proving it false — an answer showing the sample was drawn from one unusual coastline could weaken the conclusion without counting a single species. The narrow point is this: one exception logically destroys an all or none claim. One exception does not logically destroy a most claim.

So before you evaluate a single answer choice on a Strengthener or Weakener, read the conclusion and name exactly what it claims. Is it all? None? Most? Some? Probably? A comparison with something else? A cause? Merely one contributing factor? Different conclusions need different kinds and amounts of evidence to move them, and that question is far more reliable than asking whether an answer “sounds strong.”

This replaces a rule of mine that was too blunt. I have told students that an answer beginning with some is a red flag on a Strengthener or Weakener. As a rough statistical observation about what tends to appear in answer choices, that is not useless. As a rule it is wrong, and I would rather you had the real one: ask whether that amount of information is enough to move this argument. Sometimes one narrow fact is decisive. Sometimes it is irrelevant. The word at the front of the answer does not tell you which.

The Floor Test

For vague words, use what I call the Floor Test. The idea is simple:

Ask for the minimum the sentence has actually established. Do not read a stronger version into it just because the stronger version feels natural.

Example — HELPED

Stimulus: The new drainage ditch helped reduce standing water on the property.

How much did the ditch actually do? Maybe it eliminated nearly all the standing water. Maybe it was one of several major causes. Maybe it was responsible for a tiny fraction of the improvement. All three are compatible with the word helped.

So the stimulus has established only that the ditch made some contribution. An answer saying the ditch eliminated the standing-water problem says more than that. An answer saying the ditch was responsible for the reduction may also say more, if “responsible for” assigns the result to the ditch in a way “helped” did not.

Run the same check on anything:

  • some → at least one
  • may → possible, and nothing more
  • helped → made some contribution
  • up to 40 percent → a ceiling, not a number near 40
  • consistent with → compatible with the evidence, not established by it

The Floor Test stops your own brain from strengthening the author’s sentence on their behalf, which is what it will do if you let it.

The two four-letter words

If you take one habit from this article, train yourself to notice two words on sight: MOST and MORE.

MOST means more than 50 percent. Students read past it because it sounds casual in ordinary English, but on the LSAT it changes the logical burden of the sentence. And remember that it permits all: “most students passed” does not establish that anyone failed.

MORE creates a comparison, and it is one of the easiest ways for a wrong answer to smuggle in information that was never established. If a passage tells you experienced attorneys are highly effective negotiators, you cannot infer they are more effective than inexperienced ones. Maybe they are. The passage said nothing about inexperienced attorneys. That sentence did not restate the first one more confidently — it introduced a second group. The trap appears constantly on Inference, in Reading Comprehension, and on opinion questions.

What this looks like at the desk

  • Identify the mode. Is the answer choice being accepted as a new fact? If no, Mode 1. If yes, Mode 2.
  • In Mode 1, ask what the stimulus actually supports, and hold the answer to strength ≤ support.
  • In Mode 2, ask what the answer has to accomplish. Do not reach for the strongest answer by reflex. Ask whether it is strong enough to strengthen, weaken, guarantee or explain what the stem requires.
  • Read the conclusion carefully before the answers, especially on Strengthen and Weaken. Note its quantity, its certainty, any comparison, any causal claim, and its scope.
  • Apply the Floor Test to flexible words. Judge the minimum the sentence guarantees, not the stronger reading your brain supplied.
  • Watch for comparisons that arrived without a second thing to compare to. It is one of the most common unforced errors on Inference.
  • Evaluate the whole proposition, not isolated vocabulary. A forceful word inside “may,” “fails to rule out,” or “overlooks the possibility that” does not make the answer itself strong.

Language strength is not a vocabulary trick. It is the practice of tracking exactly what each statement commits you to — and refusing to read more into it than it actually says.

This is one part of a larger method. The full LSAT Masterclass covers every question type, the flaw taxonomy, the Proof Test, the Negation Test, conditional reasoning, recurring argument structures and Reading Comprehension, worked against the official tests in your own LawHub account. The rules in it are drilled in Cold Recall, the spaced-repetition trainer built from the guide. Both are included with tutoring at no extra cost.

If your score has stalled and you cannot identify exactly why you are still missing questions, that is the problem I am best at diagnosing. Book a free consultation and tell me where you are.