Flag only first row where condition is met in a DataFrameAdd one row to pandas DataFrameFilter dataframe rows if value in column is in a set list of valuesUse a list of values to select rows from a pandas dataframeHow to drop rows of Pandas DataFrame whose value in certain columns is NaNHow do I get the row count of a Pandas dataframe?Selecting a row of pandas series/dataframe by integer indexHow to iterate over rows in a DataFrame in Pandas?Select rows from a DataFrame based on values in a column in pandasDeleting DataFrame row in Pandas based on column valuer - reorder certain rows if condition is met

How do you conduct xenoanthropology after first contact?

I probably found a bug with the sudo apt install function

Why did the Germans forbid the possession of pet pigeons in Rostov-on-Don in 1941?

A function which translates a sentence to title-case

Copenhagen passport control - US citizen

Is it possible to do 50 km distance without any previous training?

Simulate Bitwise Cyclic Tag

How can the DM most effectively choose 1 out of an odd number of players to be targeted by an attack or effect?

"which" command doesn't work / path of Safari?

How do we improve the relationship with a client software team that performs poorly and is becoming less collaborative?

Possibly bubble sort algorithm

A newer friend of my brother's gave him a load of baseball cards that are supposedly extremely valuable. Is this a scam?

How can bays and straits be determined in a procedurally generated map?

Why CLRS example on residual networks does not follows its formula?

Why is this code 6.5x slower with optimizations enabled?

What are these boxed doors outside store fronts in New York?

Should I join office cleaning event for free?

I see my dog run

Why Is Death Allowed In the Matrix?

Why are 150k or 200k jobs considered good when there are 300k+ births a month?

Copycat chess is back

What defenses are there against being summoned by the Gate spell?

Draw simple lines in Inkscape

How do I create uniquely male characters?



Flag only first row where condition is met in a DataFrame


Add one row to pandas DataFrameFilter dataframe rows if value in column is in a set list of valuesUse a list of values to select rows from a pandas dataframeHow to drop rows of Pandas DataFrame whose value in certain columns is NaNHow do I get the row count of a Pandas dataframe?Selecting a row of pandas series/dataframe by integer indexHow to iterate over rows in a DataFrame in Pandas?Select rows from a DataFrame based on values in a column in pandasDeleting DataFrame row in Pandas based on column valuer - reorder certain rows if condition is met






.everyoneloves__top-leaderboard:empty,.everyoneloves__mid-leaderboard:empty,.everyoneloves__bot-mid-leaderboard:empty height:90px;width:728px;box-sizing:border-box;








8















I have the following DataFrame df, which can be created as follows:



date_today = datetime.now().date()
days = pd.date_range(date_today, date_today + timedelta(19), freq='D')
x = np.arange(0,2*np.pi,0.1*np.pi) # start,stop,step
y = np.sin(x)
df = pd.DataFrame('dates': days, 'vals': y, 'is_hit': abs(y)>0.9)
df = df.set_index('dates')


And which looks like this:



 is_hit vals
dates
2019-03-27 False 0.000000e+00
2019-03-28 False 3.090170e-01
2019-03-29 False 5.877853e-01
2019-03-30 False 8.090170e-01
2019-03-31 True 9.510565e-01
2019-04-01 True 1.000000e+00
2019-04-02 True 9.510565e-01
2019-04-03 False 8.090170e-01
2019-04-04 False 5.877853e-01
2019-04-05 False 3.090170e-01
2019-04-06 False 1.224647e-16
2019-04-07 False -3.090170e-01
2019-04-08 False -5.877853e-01
2019-04-09 False -8.090170e-01
2019-04-10 True -9.510565e-01
2019-04-11 True -1.000000e+00
2019-04-12 True -9.510565e-01
2019-04-13 False -8.090170e-01
2019-04-14 False -5.877853e-01
2019-04-15 False -3.090170e-01


I want to flag the rows where the is_hit condition is True for the first time, such that the expected new column hit_first would be:



 is_hit vals hit_first
dates
2019-03-27 False 0.000000e+00 False
2019-03-28 False 3.090170e-01 False
2019-03-29 False 5.877853e-01 False
2019-03-30 False 8.090170e-01 False
2019-03-31 True 9.510565e-01 True
2019-04-01 True 1.000000e+00 False
2019-04-02 True 9.510565e-01 False
2019-04-03 False 8.090170e-01 False
2019-04-04 False 5.877853e-01 False
2019-04-05 False 3.090170e-01 False
2019-04-06 False 1.224647e-16 False
2019-04-07 False -3.090170e-01 False
2019-04-08 False -5.877853e-01 False
2019-04-09 False -8.090170e-01 False
2019-04-10 True -9.510565e-01 True
2019-04-11 True -1.000000e+00 False
2019-04-12 True -9.510565e-01 False
2019-04-13 False -8.090170e-01 False
2019-04-14 False -5.877853e-01 False
2019-04-15 False -3.090170e-01 False









share|improve this question




























    8















    I have the following DataFrame df, which can be created as follows:



    date_today = datetime.now().date()
    days = pd.date_range(date_today, date_today + timedelta(19), freq='D')
    x = np.arange(0,2*np.pi,0.1*np.pi) # start,stop,step
    y = np.sin(x)
    df = pd.DataFrame('dates': days, 'vals': y, 'is_hit': abs(y)>0.9)
    df = df.set_index('dates')


    And which looks like this:



     is_hit vals
    dates
    2019-03-27 False 0.000000e+00
    2019-03-28 False 3.090170e-01
    2019-03-29 False 5.877853e-01
    2019-03-30 False 8.090170e-01
    2019-03-31 True 9.510565e-01
    2019-04-01 True 1.000000e+00
    2019-04-02 True 9.510565e-01
    2019-04-03 False 8.090170e-01
    2019-04-04 False 5.877853e-01
    2019-04-05 False 3.090170e-01
    2019-04-06 False 1.224647e-16
    2019-04-07 False -3.090170e-01
    2019-04-08 False -5.877853e-01
    2019-04-09 False -8.090170e-01
    2019-04-10 True -9.510565e-01
    2019-04-11 True -1.000000e+00
    2019-04-12 True -9.510565e-01
    2019-04-13 False -8.090170e-01
    2019-04-14 False -5.877853e-01
    2019-04-15 False -3.090170e-01


    I want to flag the rows where the is_hit condition is True for the first time, such that the expected new column hit_first would be:



     is_hit vals hit_first
    dates
    2019-03-27 False 0.000000e+00 False
    2019-03-28 False 3.090170e-01 False
    2019-03-29 False 5.877853e-01 False
    2019-03-30 False 8.090170e-01 False
    2019-03-31 True 9.510565e-01 True
    2019-04-01 True 1.000000e+00 False
    2019-04-02 True 9.510565e-01 False
    2019-04-03 False 8.090170e-01 False
    2019-04-04 False 5.877853e-01 False
    2019-04-05 False 3.090170e-01 False
    2019-04-06 False 1.224647e-16 False
    2019-04-07 False -3.090170e-01 False
    2019-04-08 False -5.877853e-01 False
    2019-04-09 False -8.090170e-01 False
    2019-04-10 True -9.510565e-01 True
    2019-04-11 True -1.000000e+00 False
    2019-04-12 True -9.510565e-01 False
    2019-04-13 False -8.090170e-01 False
    2019-04-14 False -5.877853e-01 False
    2019-04-15 False -3.090170e-01 False









    share|improve this question
























      8












      8








      8








      I have the following DataFrame df, which can be created as follows:



      date_today = datetime.now().date()
      days = pd.date_range(date_today, date_today + timedelta(19), freq='D')
      x = np.arange(0,2*np.pi,0.1*np.pi) # start,stop,step
      y = np.sin(x)
      df = pd.DataFrame('dates': days, 'vals': y, 'is_hit': abs(y)>0.9)
      df = df.set_index('dates')


      And which looks like this:



       is_hit vals
      dates
      2019-03-27 False 0.000000e+00
      2019-03-28 False 3.090170e-01
      2019-03-29 False 5.877853e-01
      2019-03-30 False 8.090170e-01
      2019-03-31 True 9.510565e-01
      2019-04-01 True 1.000000e+00
      2019-04-02 True 9.510565e-01
      2019-04-03 False 8.090170e-01
      2019-04-04 False 5.877853e-01
      2019-04-05 False 3.090170e-01
      2019-04-06 False 1.224647e-16
      2019-04-07 False -3.090170e-01
      2019-04-08 False -5.877853e-01
      2019-04-09 False -8.090170e-01
      2019-04-10 True -9.510565e-01
      2019-04-11 True -1.000000e+00
      2019-04-12 True -9.510565e-01
      2019-04-13 False -8.090170e-01
      2019-04-14 False -5.877853e-01
      2019-04-15 False -3.090170e-01


      I want to flag the rows where the is_hit condition is True for the first time, such that the expected new column hit_first would be:



       is_hit vals hit_first
      dates
      2019-03-27 False 0.000000e+00 False
      2019-03-28 False 3.090170e-01 False
      2019-03-29 False 5.877853e-01 False
      2019-03-30 False 8.090170e-01 False
      2019-03-31 True 9.510565e-01 True
      2019-04-01 True 1.000000e+00 False
      2019-04-02 True 9.510565e-01 False
      2019-04-03 False 8.090170e-01 False
      2019-04-04 False 5.877853e-01 False
      2019-04-05 False 3.090170e-01 False
      2019-04-06 False 1.224647e-16 False
      2019-04-07 False -3.090170e-01 False
      2019-04-08 False -5.877853e-01 False
      2019-04-09 False -8.090170e-01 False
      2019-04-10 True -9.510565e-01 True
      2019-04-11 True -1.000000e+00 False
      2019-04-12 True -9.510565e-01 False
      2019-04-13 False -8.090170e-01 False
      2019-04-14 False -5.877853e-01 False
      2019-04-15 False -3.090170e-01 False









      share|improve this question














      I have the following DataFrame df, which can be created as follows:



      date_today = datetime.now().date()
      days = pd.date_range(date_today, date_today + timedelta(19), freq='D')
      x = np.arange(0,2*np.pi,0.1*np.pi) # start,stop,step
      y = np.sin(x)
      df = pd.DataFrame('dates': days, 'vals': y, 'is_hit': abs(y)>0.9)
      df = df.set_index('dates')


      And which looks like this:



       is_hit vals
      dates
      2019-03-27 False 0.000000e+00
      2019-03-28 False 3.090170e-01
      2019-03-29 False 5.877853e-01
      2019-03-30 False 8.090170e-01
      2019-03-31 True 9.510565e-01
      2019-04-01 True 1.000000e+00
      2019-04-02 True 9.510565e-01
      2019-04-03 False 8.090170e-01
      2019-04-04 False 5.877853e-01
      2019-04-05 False 3.090170e-01
      2019-04-06 False 1.224647e-16
      2019-04-07 False -3.090170e-01
      2019-04-08 False -5.877853e-01
      2019-04-09 False -8.090170e-01
      2019-04-10 True -9.510565e-01
      2019-04-11 True -1.000000e+00
      2019-04-12 True -9.510565e-01
      2019-04-13 False -8.090170e-01
      2019-04-14 False -5.877853e-01
      2019-04-15 False -3.090170e-01


      I want to flag the rows where the is_hit condition is True for the first time, such that the expected new column hit_first would be:



       is_hit vals hit_first
      dates
      2019-03-27 False 0.000000e+00 False
      2019-03-28 False 3.090170e-01 False
      2019-03-29 False 5.877853e-01 False
      2019-03-30 False 8.090170e-01 False
      2019-03-31 True 9.510565e-01 True
      2019-04-01 True 1.000000e+00 False
      2019-04-02 True 9.510565e-01 False
      2019-04-03 False 8.090170e-01 False
      2019-04-04 False 5.877853e-01 False
      2019-04-05 False 3.090170e-01 False
      2019-04-06 False 1.224647e-16 False
      2019-04-07 False -3.090170e-01 False
      2019-04-08 False -5.877853e-01 False
      2019-04-09 False -8.090170e-01 False
      2019-04-10 True -9.510565e-01 True
      2019-04-11 True -1.000000e+00 False
      2019-04-12 True -9.510565e-01 False
      2019-04-13 False -8.090170e-01 False
      2019-04-14 False -5.877853e-01 False
      2019-04-15 False -3.090170e-01 False






      python pandas dataframe






      share|improve this question













      share|improve this question











      share|improve this question




      share|improve this question










      asked Mar 27 at 12:23









      JejeBelfortJejeBelfort

      6911624




      6911624






















          4 Answers
          4






          active

          oldest

          votes


















          10














          My suggestion:



          df['hit_first'] = df['is_hit'] & (~df['is_hit']).shift(1)





          share|improve this answer






























            3














            Use Series.shift chained with & for bitwise AND:



            df['hit_first'] = df['is_hit'].ne(df['is_hit'].shift()) & df['is_hit']
            print (df)
            vals is_hit hit_first
            dates
            2019-03-27 0.000000e+00 False False
            2019-03-28 3.090170e-01 False False
            2019-03-29 5.877853e-01 False False
            2019-03-30 8.090170e-01 False False
            2019-03-31 9.510565e-01 True True
            2019-04-01 1.000000e+00 True False
            2019-04-02 9.510565e-01 True False
            2019-04-03 8.090170e-01 False False
            2019-04-04 5.877853e-01 False False
            2019-04-05 3.090170e-01 False False
            2019-04-06 1.224647e-16 False False
            2019-04-07 -3.090170e-01 False False
            2019-04-08 -5.877853e-01 False False
            2019-04-09 -8.090170e-01 False False
            2019-04-10 -9.510565e-01 True True
            2019-04-11 -1.000000e+00 True False
            2019-04-12 -9.510565e-01 True False
            2019-04-13 -8.090170e-01 False False
            2019-04-14 -5.877853e-01 False False
            2019-04-15 -3.090170e-01 False False





            share|improve this answer
































              3














              I also, think you can do it this way:



              df['is_hit'].astype(int).diff() == 1


              Output:



              dates
              2019-03-27 False
              2019-03-28 False
              2019-03-29 False
              2019-03-30 False
              2019-03-31 True
              2019-04-01 False
              2019-04-02 False
              2019-04-03 False
              2019-04-04 False
              2019-04-05 False
              2019-04-06 False
              2019-04-07 False
              2019-04-08 False
              2019-04-09 False
              2019-04-10 True
              2019-04-11 False
              2019-04-12 False
              2019-04-13 False
              2019-04-14 False
              2019-04-15 False
              Name: is_hit, dtype: bool


              Timings:



              %timeit df['is_hit'] & (~df['is_hit']).shift(1)
              1.13 ms ± 5.63 µs per loop (mean ± std. dev. of 7 runs, 1000 loops each)

              %timeit df['is_hit'].ne(df['is_hit'].shift()) & df['is_hit']
              908 µs ± 9.53 µs per loop (mean ± std. dev. of 7 runs, 1000 loops each)

              %timeit df['is_hit'].astype(int).diff() == 1
              689 µs ± 8.24 µs per loop (mean ± std. dev. of 7 runs, 1000 loops each)





              share|improve this answer




















              • 2





                Nice, maybe performance in large data should be interesting.

                – jezrael
                Mar 27 at 13:11



















              -1














              Also this can be done by using simple difference between the series and it's shifted series by 1 period :



              df['hit_first'] = df['is_hit']-df['is_hit'].shift()==1





              share|improve this answer




















              • 1





                The use of np.where here is quite pointless.

                – miradulo
                Mar 27 at 18:50











              • Yes I understood. Thanks :)

                – Loochie
                Mar 27 at 20:42











              • While this code may answer the question, providing additional context regarding how and/or why it solves the problem would improve the answer's long-term value.

                – DebanjanB
                Mar 27 at 21:22











              Your Answer






              StackExchange.ifUsing("editor", function ()
              StackExchange.using("externalEditor", function ()
              StackExchange.using("snippets", function ()
              StackExchange.snippets.init();
              );
              );
              , "code-snippets");

              StackExchange.ready(function()
              var channelOptions =
              tags: "".split(" "),
              id: "1"
              ;
              initTagRenderer("".split(" "), "".split(" "), channelOptions);

              StackExchange.using("externalEditor", function()
              // Have to fire editor after snippets, if snippets enabled
              if (StackExchange.settings.snippets.snippetsEnabled)
              StackExchange.using("snippets", function()
              createEditor();
              );

              else
              createEditor();

              );

              function createEditor()
              StackExchange.prepareEditor(
              heartbeatType: 'answer',
              autoActivateHeartbeat: false,
              convertImagesToLinks: true,
              noModals: true,
              showLowRepImageUploadWarning: true,
              reputationToPostImages: 10,
              bindNavPrevention: true,
              postfix: "",
              imageUploader:
              brandingHtml: "Powered by u003ca class="icon-imgur-white" href="https://imgur.com/"u003eu003c/au003e",
              contentPolicyHtml: "User contributions licensed under u003ca href="https://creativecommons.org/licenses/by-sa/3.0/"u003ecc by-sa 3.0 with attribution requiredu003c/au003e u003ca href="https://stackoverflow.com/legal/content-policy"u003e(content policy)u003c/au003e",
              allowUrls: true
              ,
              onDemand: true,
              discardSelector: ".discard-answer"
              ,immediatelyShowMarkdownHelp:true
              );



              );













              draft saved

              draft discarded


















              StackExchange.ready(
              function ()
              StackExchange.openid.initPostLogin('.new-post-login', 'https%3a%2f%2fstackoverflow.com%2fquestions%2f55377130%2fflag-only-first-row-where-condition-is-met-in-a-dataframe%23new-answer', 'question_page');

              );

              Post as a guest















              Required, but never shown

























              4 Answers
              4






              active

              oldest

              votes








              4 Answers
              4






              active

              oldest

              votes









              active

              oldest

              votes






              active

              oldest

              votes









              10














              My suggestion:



              df['hit_first'] = df['is_hit'] & (~df['is_hit']).shift(1)





              share|improve this answer



























                10














                My suggestion:



                df['hit_first'] = df['is_hit'] & (~df['is_hit']).shift(1)





                share|improve this answer

























                  10












                  10








                  10







                  My suggestion:



                  df['hit_first'] = df['is_hit'] & (~df['is_hit']).shift(1)





                  share|improve this answer













                  My suggestion:



                  df['hit_first'] = df['is_hit'] & (~df['is_hit']).shift(1)






                  share|improve this answer












                  share|improve this answer



                  share|improve this answer










                  answered Mar 27 at 12:28









                  ecortazarecortazar

                  96618




                  96618























                      3














                      Use Series.shift chained with & for bitwise AND:



                      df['hit_first'] = df['is_hit'].ne(df['is_hit'].shift()) & df['is_hit']
                      print (df)
                      vals is_hit hit_first
                      dates
                      2019-03-27 0.000000e+00 False False
                      2019-03-28 3.090170e-01 False False
                      2019-03-29 5.877853e-01 False False
                      2019-03-30 8.090170e-01 False False
                      2019-03-31 9.510565e-01 True True
                      2019-04-01 1.000000e+00 True False
                      2019-04-02 9.510565e-01 True False
                      2019-04-03 8.090170e-01 False False
                      2019-04-04 5.877853e-01 False False
                      2019-04-05 3.090170e-01 False False
                      2019-04-06 1.224647e-16 False False
                      2019-04-07 -3.090170e-01 False False
                      2019-04-08 -5.877853e-01 False False
                      2019-04-09 -8.090170e-01 False False
                      2019-04-10 -9.510565e-01 True True
                      2019-04-11 -1.000000e+00 True False
                      2019-04-12 -9.510565e-01 True False
                      2019-04-13 -8.090170e-01 False False
                      2019-04-14 -5.877853e-01 False False
                      2019-04-15 -3.090170e-01 False False





                      share|improve this answer





























                        3














                        Use Series.shift chained with & for bitwise AND:



                        df['hit_first'] = df['is_hit'].ne(df['is_hit'].shift()) & df['is_hit']
                        print (df)
                        vals is_hit hit_first
                        dates
                        2019-03-27 0.000000e+00 False False
                        2019-03-28 3.090170e-01 False False
                        2019-03-29 5.877853e-01 False False
                        2019-03-30 8.090170e-01 False False
                        2019-03-31 9.510565e-01 True True
                        2019-04-01 1.000000e+00 True False
                        2019-04-02 9.510565e-01 True False
                        2019-04-03 8.090170e-01 False False
                        2019-04-04 5.877853e-01 False False
                        2019-04-05 3.090170e-01 False False
                        2019-04-06 1.224647e-16 False False
                        2019-04-07 -3.090170e-01 False False
                        2019-04-08 -5.877853e-01 False False
                        2019-04-09 -8.090170e-01 False False
                        2019-04-10 -9.510565e-01 True True
                        2019-04-11 -1.000000e+00 True False
                        2019-04-12 -9.510565e-01 True False
                        2019-04-13 -8.090170e-01 False False
                        2019-04-14 -5.877853e-01 False False
                        2019-04-15 -3.090170e-01 False False





                        share|improve this answer



























                          3












                          3








                          3







                          Use Series.shift chained with & for bitwise AND:



                          df['hit_first'] = df['is_hit'].ne(df['is_hit'].shift()) & df['is_hit']
                          print (df)
                          vals is_hit hit_first
                          dates
                          2019-03-27 0.000000e+00 False False
                          2019-03-28 3.090170e-01 False False
                          2019-03-29 5.877853e-01 False False
                          2019-03-30 8.090170e-01 False False
                          2019-03-31 9.510565e-01 True True
                          2019-04-01 1.000000e+00 True False
                          2019-04-02 9.510565e-01 True False
                          2019-04-03 8.090170e-01 False False
                          2019-04-04 5.877853e-01 False False
                          2019-04-05 3.090170e-01 False False
                          2019-04-06 1.224647e-16 False False
                          2019-04-07 -3.090170e-01 False False
                          2019-04-08 -5.877853e-01 False False
                          2019-04-09 -8.090170e-01 False False
                          2019-04-10 -9.510565e-01 True True
                          2019-04-11 -1.000000e+00 True False
                          2019-04-12 -9.510565e-01 True False
                          2019-04-13 -8.090170e-01 False False
                          2019-04-14 -5.877853e-01 False False
                          2019-04-15 -3.090170e-01 False False





                          share|improve this answer















                          Use Series.shift chained with & for bitwise AND:



                          df['hit_first'] = df['is_hit'].ne(df['is_hit'].shift()) & df['is_hit']
                          print (df)
                          vals is_hit hit_first
                          dates
                          2019-03-27 0.000000e+00 False False
                          2019-03-28 3.090170e-01 False False
                          2019-03-29 5.877853e-01 False False
                          2019-03-30 8.090170e-01 False False
                          2019-03-31 9.510565e-01 True True
                          2019-04-01 1.000000e+00 True False
                          2019-04-02 9.510565e-01 True False
                          2019-04-03 8.090170e-01 False False
                          2019-04-04 5.877853e-01 False False
                          2019-04-05 3.090170e-01 False False
                          2019-04-06 1.224647e-16 False False
                          2019-04-07 -3.090170e-01 False False
                          2019-04-08 -5.877853e-01 False False
                          2019-04-09 -8.090170e-01 False False
                          2019-04-10 -9.510565e-01 True True
                          2019-04-11 -1.000000e+00 True False
                          2019-04-12 -9.510565e-01 True False
                          2019-04-13 -8.090170e-01 False False
                          2019-04-14 -5.877853e-01 False False
                          2019-04-15 -3.090170e-01 False False






                          share|improve this answer














                          share|improve this answer



                          share|improve this answer








                          edited Mar 27 at 12:33

























                          answered Mar 27 at 12:28









                          jezraeljezrael

                          356k26320396




                          356k26320396





















                              3














                              I also, think you can do it this way:



                              df['is_hit'].astype(int).diff() == 1


                              Output:



                              dates
                              2019-03-27 False
                              2019-03-28 False
                              2019-03-29 False
                              2019-03-30 False
                              2019-03-31 True
                              2019-04-01 False
                              2019-04-02 False
                              2019-04-03 False
                              2019-04-04 False
                              2019-04-05 False
                              2019-04-06 False
                              2019-04-07 False
                              2019-04-08 False
                              2019-04-09 False
                              2019-04-10 True
                              2019-04-11 False
                              2019-04-12 False
                              2019-04-13 False
                              2019-04-14 False
                              2019-04-15 False
                              Name: is_hit, dtype: bool


                              Timings:



                              %timeit df['is_hit'] & (~df['is_hit']).shift(1)
                              1.13 ms ± 5.63 µs per loop (mean ± std. dev. of 7 runs, 1000 loops each)

                              %timeit df['is_hit'].ne(df['is_hit'].shift()) & df['is_hit']
                              908 µs ± 9.53 µs per loop (mean ± std. dev. of 7 runs, 1000 loops each)

                              %timeit df['is_hit'].astype(int).diff() == 1
                              689 µs ± 8.24 µs per loop (mean ± std. dev. of 7 runs, 1000 loops each)





                              share|improve this answer




















                              • 2





                                Nice, maybe performance in large data should be interesting.

                                – jezrael
                                Mar 27 at 13:11
















                              3














                              I also, think you can do it this way:



                              df['is_hit'].astype(int).diff() == 1


                              Output:



                              dates
                              2019-03-27 False
                              2019-03-28 False
                              2019-03-29 False
                              2019-03-30 False
                              2019-03-31 True
                              2019-04-01 False
                              2019-04-02 False
                              2019-04-03 False
                              2019-04-04 False
                              2019-04-05 False
                              2019-04-06 False
                              2019-04-07 False
                              2019-04-08 False
                              2019-04-09 False
                              2019-04-10 True
                              2019-04-11 False
                              2019-04-12 False
                              2019-04-13 False
                              2019-04-14 False
                              2019-04-15 False
                              Name: is_hit, dtype: bool


                              Timings:



                              %timeit df['is_hit'] & (~df['is_hit']).shift(1)
                              1.13 ms ± 5.63 µs per loop (mean ± std. dev. of 7 runs, 1000 loops each)

                              %timeit df['is_hit'].ne(df['is_hit'].shift()) & df['is_hit']
                              908 µs ± 9.53 µs per loop (mean ± std. dev. of 7 runs, 1000 loops each)

                              %timeit df['is_hit'].astype(int).diff() == 1
                              689 µs ± 8.24 µs per loop (mean ± std. dev. of 7 runs, 1000 loops each)





                              share|improve this answer




















                              • 2





                                Nice, maybe performance in large data should be interesting.

                                – jezrael
                                Mar 27 at 13:11














                              3












                              3








                              3







                              I also, think you can do it this way:



                              df['is_hit'].astype(int).diff() == 1


                              Output:



                              dates
                              2019-03-27 False
                              2019-03-28 False
                              2019-03-29 False
                              2019-03-30 False
                              2019-03-31 True
                              2019-04-01 False
                              2019-04-02 False
                              2019-04-03 False
                              2019-04-04 False
                              2019-04-05 False
                              2019-04-06 False
                              2019-04-07 False
                              2019-04-08 False
                              2019-04-09 False
                              2019-04-10 True
                              2019-04-11 False
                              2019-04-12 False
                              2019-04-13 False
                              2019-04-14 False
                              2019-04-15 False
                              Name: is_hit, dtype: bool


                              Timings:



                              %timeit df['is_hit'] & (~df['is_hit']).shift(1)
                              1.13 ms ± 5.63 µs per loop (mean ± std. dev. of 7 runs, 1000 loops each)

                              %timeit df['is_hit'].ne(df['is_hit'].shift()) & df['is_hit']
                              908 µs ± 9.53 µs per loop (mean ± std. dev. of 7 runs, 1000 loops each)

                              %timeit df['is_hit'].astype(int).diff() == 1
                              689 µs ± 8.24 µs per loop (mean ± std. dev. of 7 runs, 1000 loops each)





                              share|improve this answer















                              I also, think you can do it this way:



                              df['is_hit'].astype(int).diff() == 1


                              Output:



                              dates
                              2019-03-27 False
                              2019-03-28 False
                              2019-03-29 False
                              2019-03-30 False
                              2019-03-31 True
                              2019-04-01 False
                              2019-04-02 False
                              2019-04-03 False
                              2019-04-04 False
                              2019-04-05 False
                              2019-04-06 False
                              2019-04-07 False
                              2019-04-08 False
                              2019-04-09 False
                              2019-04-10 True
                              2019-04-11 False
                              2019-04-12 False
                              2019-04-13 False
                              2019-04-14 False
                              2019-04-15 False
                              Name: is_hit, dtype: bool


                              Timings:



                              %timeit df['is_hit'] & (~df['is_hit']).shift(1)
                              1.13 ms ± 5.63 µs per loop (mean ± std. dev. of 7 runs, 1000 loops each)

                              %timeit df['is_hit'].ne(df['is_hit'].shift()) & df['is_hit']
                              908 µs ± 9.53 µs per loop (mean ± std. dev. of 7 runs, 1000 loops each)

                              %timeit df['is_hit'].astype(int).diff() == 1
                              689 µs ± 8.24 µs per loop (mean ± std. dev. of 7 runs, 1000 loops each)






                              share|improve this answer














                              share|improve this answer



                              share|improve this answer








                              edited Mar 27 at 12:57

























                              answered Mar 27 at 12:50









                              Scott BostonScott Boston

                              58k73258




                              58k73258







                              • 2





                                Nice, maybe performance in large data should be interesting.

                                – jezrael
                                Mar 27 at 13:11













                              • 2





                                Nice, maybe performance in large data should be interesting.

                                – jezrael
                                Mar 27 at 13:11








                              2




                              2





                              Nice, maybe performance in large data should be interesting.

                              – jezrael
                              Mar 27 at 13:11






                              Nice, maybe performance in large data should be interesting.

                              – jezrael
                              Mar 27 at 13:11












                              -1














                              Also this can be done by using simple difference between the series and it's shifted series by 1 period :



                              df['hit_first'] = df['is_hit']-df['is_hit'].shift()==1





                              share|improve this answer




















                              • 1





                                The use of np.where here is quite pointless.

                                – miradulo
                                Mar 27 at 18:50











                              • Yes I understood. Thanks :)

                                – Loochie
                                Mar 27 at 20:42











                              • While this code may answer the question, providing additional context regarding how and/or why it solves the problem would improve the answer's long-term value.

                                – DebanjanB
                                Mar 27 at 21:22















                              -1














                              Also this can be done by using simple difference between the series and it's shifted series by 1 period :



                              df['hit_first'] = df['is_hit']-df['is_hit'].shift()==1





                              share|improve this answer




















                              • 1





                                The use of np.where here is quite pointless.

                                – miradulo
                                Mar 27 at 18:50











                              • Yes I understood. Thanks :)

                                – Loochie
                                Mar 27 at 20:42











                              • While this code may answer the question, providing additional context regarding how and/or why it solves the problem would improve the answer's long-term value.

                                – DebanjanB
                                Mar 27 at 21:22













                              -1












                              -1








                              -1







                              Also this can be done by using simple difference between the series and it's shifted series by 1 period :



                              df['hit_first'] = df['is_hit']-df['is_hit'].shift()==1





                              share|improve this answer















                              Also this can be done by using simple difference between the series and it's shifted series by 1 period :



                              df['hit_first'] = df['is_hit']-df['is_hit'].shift()==1






                              share|improve this answer














                              share|improve this answer



                              share|improve this answer








                              edited Mar 27 at 22:45

























                              answered Mar 27 at 12:54









                              LoochieLoochie

                              959310




                              959310







                              • 1





                                The use of np.where here is quite pointless.

                                – miradulo
                                Mar 27 at 18:50











                              • Yes I understood. Thanks :)

                                – Loochie
                                Mar 27 at 20:42











                              • While this code may answer the question, providing additional context regarding how and/or why it solves the problem would improve the answer's long-term value.

                                – DebanjanB
                                Mar 27 at 21:22












                              • 1





                                The use of np.where here is quite pointless.

                                – miradulo
                                Mar 27 at 18:50











                              • Yes I understood. Thanks :)

                                – Loochie
                                Mar 27 at 20:42











                              • While this code may answer the question, providing additional context regarding how and/or why it solves the problem would improve the answer's long-term value.

                                – DebanjanB
                                Mar 27 at 21:22







                              1




                              1





                              The use of np.where here is quite pointless.

                              – miradulo
                              Mar 27 at 18:50





                              The use of np.where here is quite pointless.

                              – miradulo
                              Mar 27 at 18:50













                              Yes I understood. Thanks :)

                              – Loochie
                              Mar 27 at 20:42





                              Yes I understood. Thanks :)

                              – Loochie
                              Mar 27 at 20:42













                              While this code may answer the question, providing additional context regarding how and/or why it solves the problem would improve the answer's long-term value.

                              – DebanjanB
                              Mar 27 at 21:22





                              While this code may answer the question, providing additional context regarding how and/or why it solves the problem would improve the answer's long-term value.

                              – DebanjanB
                              Mar 27 at 21:22

















                              draft saved

                              draft discarded
















































                              Thanks for contributing an answer to Stack Overflow!


                              • Please be sure to answer the question. Provide details and share your research!

                              But avoid …


                              • Asking for help, clarification, or responding to other answers.

                              • Making statements based on opinion; back them up with references or personal experience.

                              To learn more, see our tips on writing great answers.




                              draft saved


                              draft discarded














                              StackExchange.ready(
                              function ()
                              StackExchange.openid.initPostLogin('.new-post-login', 'https%3a%2f%2fstackoverflow.com%2fquestions%2f55377130%2fflag-only-first-row-where-condition-is-met-in-a-dataframe%23new-answer', 'question_page');

                              );

                              Post as a guest















                              Required, but never shown





















































                              Required, but never shown














                              Required, but never shown












                              Required, but never shown







                              Required, but never shown

































                              Required, but never shown














                              Required, but never shown












                              Required, but never shown







                              Required, but never shown







                              -dataframe, pandas, python

                              Popular posts from this blog

                              Word for a person who has no opinion about whether god existsWord for having a definite opinion while simultaneously withholding judgment?What's the opposite of “newcomer? Is ”veteran" OK?What do you call an “atheist” who might believe in an afterlife?What's a word for someone who wants to voice opinions but not have them challenged?Word for someone who dismisses contrary opinions as irrational?Somone who thinks they are overly special/out of the ordinaryIs there a word, phrase or idiom for “a person who is incapable of thinking about the future”?The belief that a god is human-likeA word for a non-famous person/thing you have heard a lot aboutAdjective for a person who enjoys taking care of their appearance

                              2017 IndyCar Series Contents Series news Teams and drivers Schedule Season summary Footnotes References External links Navigation menu"INDYCAR: Initial 2018 bodywork concepts unveiled"the original"IndyCar confirms switch to Performance Friction brakes in 2017""AJ Foyt Racing will switch to Chevy"the original"Carlos Munoz, Conor Daly will drive for AJ Foyt Racing""Zach Veach's Indy 500 Debut Confirmed with Foyt""No mass exodus from Honda after Ganassi switch""Ex-F1 driver Sato joins Andretti Autosport for 2017 IndyCar season""IndyCar's Ryan Hunter-Reay, sponsor DHL paired through 2020""hhgregg and Andretti Autosport announce partnership for key races in 2016""INDYCAR: Rossi re-signs with Andretti"the original"McLaren Formula 1 - Fernando Alonso to race at Indy 500 with McLaren, Honda and Andretti Autosport""Shank will finally take part in Indy 500 with Harvey, Andretti | MotorSportsTalk""Andretti adds Jack Harvey to Indy 500 field""Ganassi switches to Honda power for 2017""INDYCAR: Chilton returns to Ganassi"the original"IndyCar silly season: Who's going where in 2017?""INDYCAR: Kanaan, NTT Data return to Ganassi"the original"Kimball to remain at Ganassi for 2017""Coyne confirms Bourdais for 2017 IndyCar season""Davison to sub for Bourdais in Indy 500"the original"Gutierrez confirmed for Detroit IndyCar debut""Gutierrez returns with Coyne for rest of 2017 season""Vautier to drive for Coyne at Texas"the original"INDYCAR: Coyne confirms Jones for 2017"the original"Pippa Mann returns to Coyne for Indy 500""Karam, Dreyer & Reinbold teaming up again for Indianapolis 500""Pigot to return to Ed Carpenter Racing""Hildebrand confirmed as full-time Ed Carpenter driver""Veach to replace injured Hildebrand at Barber"the originalNew Team Harding Racing Enters Chaves for 101st Indianapolis 500"Juncos Racing Announces Entry in 101st Running of the Indianapolis 500 :: Juncos Racing""Juncos confirms Pigot for Indy 500""Saavedra confirmed in Juncos' second 500 entry"the original"Lazier confirms Indy 500 run after son's USF2000 debut"the original"Claman DeMelo to race for RLLR at Sonoma"the original"Rahal signs Servia and ace engineer for 2017""IndyCar: Aleshin returns with Schmidt"the original"Aleshin replaced by Saavedra for Toronto""Jack Harvey will pilot SPM No. 7 car at Watkins Glen, Sonoma""Jay Howard confirmed in Tony Stewart's supported SPM Indy entry""INDYCAR: Newgarden to wave the flag at Penske"the original"Pagenaud opts for No. 1 in 2017"the original"Penske confirms Newgarden for 2017""Montoya to stay with Team Penske in 2017""Target leaving IndyCar after 27 seasons with Chip Ganassi""Cavin: IndyCar could see complete driver/team shakeup in 2017""End of the road for KV Racing?""KV Racing confirms closure, equipment sold to Juncos""Juncos confirms IndyCar Series entry"the original"Juncos readies IndyCar program, aims for '17 500"the original"Harding Racing to add Texas, Pocono to schedule"the original"Sato signs with Andretti Autosport for 2017""INDYCAR: Aleshin in Doubt at SPM"the original"Long Beach notebook: JR Hildebrand breaks hand""Hildebrand cleared to return at Phoenix"the original"Bourdais to undergo surgery on multiple fractures""Aleshin loses Schmidt Peterson IndyCar ride""Saavedra in at SPM for Pocono, Gateway"the original"Bourdais to make return at Gateway"the original"The IndyCar Grand Prix no longer is sponsored by Angie's List""2017 IndyCar Series rulebook""2017 Verizon IndyCar Series Official Rulebook"Official websiteeeeee

                              Can I sign legal documents with a smiley face?Do Legal Documents Require Signing In Standard Pen Colors?Is it possible to legally prohibit someone from linking to specific pages on your website?Do scans of signed documents have the same legal power as the original document?What can I do if I signed an excessively restrictive contract?Can other party sneak in new contract terms via termination notice?How to prove that someone forged my signature on a contract that I was not aware of?In Australia, Is it legal to sign a document as somebody else?making a contract that includes video licenceLease dispute, over email and text messageIf you must include all of the natural language prose in a legal document, or if it can be abstracted outE-signing: legal ramifications of “identifying” a person